🤗 Transformers: the model-definition framework for state-of-the-art machine learning models in text, vision, audio, and multimodal models, for both inference and training.
Annuaire de skills
Découvrez des skills réutilisables pour les AI agents.
Chaque recommandation reste clairement reliée à son dépôt, son audit et son chemin d’installation.
Résultats de recherche: efficient-inference
Annuaire en anglaisMNN: A blazing-fast, lightweight inference engine battle-tested by Alibaba, powering high-performance on-device LLMs and Edge AI.
DeepSpeed is a deep learning optimization library that makes distributed training and inference easy, efficient, and effective.
Multi-lingual large voice generation model, providing inference, training and deployment full-stack ability.
🤗 The largest hub of ready-to-use datasets for AI models with fast, easy-to-use and efficient data manipulation tools
An Industrial-Level Controllable and Efficient Zero-Shot Text-To-Speech System
Run agents like Hermes and OpenClaw more securely inside NVIDIA OpenShell with managed inference
A powerful framework for faster, easier, and more efficient project development.
Open source alternative to AWS. Elastic compute, block storage (non replicated), firewall and load balancer, managed Postgres, K8s, AI inference, and IAM services.
The Triton Inference Server provides an optimized cloud and edge inferencing solution.
OpenVINO™ is an open source toolkit for optimizing and deploying AI inference
An alternative privacy-friendly YouTube frontend which is efficient by design.