Directorio de skills

Descubre skills reutilizables para AI agents.

Busca skills reales de GitHub por tarea y revisa stars, confianza, auditoría, categoría y ruta de instalación antes de utilizarlos.

Cada recomendación conserva un vínculo claro con su repositorio, auditoría y ruta de instalación.

Resultados de búsqueda: inference-acceleration

Directorio en inglés

Tensors and Dynamic neural networks in Python with strong GPU acceleration

101K
Stars
76/100
Confianza
Categoría: ml-automationAuditoría

🤗 Transformers: the model-definition framework for state-of-the-art machine learning models in text, vision, audio, and multimodal models, for both inference and training.

162K
Stars
87/100
Confianza
Categoría: ml-automationAuditoría
MNN88

MNN: A blazing-fast, lightweight inference engine battle-tested by Alibaba, powering high-performance on-device LLMs and Edge AI.

16K
Stars
88/100
Confianza
Categoría: ml-automationAuditoría

DeepSpeed is a deep learning optimization library that makes distributed training and inference easy, efficient, and effective.

43K
Stars
87/100
Confianza
Categoría: ml-automationAuditoría

Multi-lingual large voice generation model, providing inference, training and deployment full-stack ability.

22K
Stars
87/100
Confianza
Categoría: support-automationAuditoría

Run agents like Hermes and OpenClaw more securely inside NVIDIA OpenShell with managed inference

21K
Stars
87/100
Confianza
Categoría: agent-frameworksAuditoría

Open source alternative to AWS. Elastic compute, block storage (non replicated), firewall and load balancer, managed Postgres, K8s, AI inference, and IAM services.

12K
Stars
85/100
Confianza
Categoría: github-automationAuditoría

The Triton Inference Server provides an optimized cloud and edge inferencing solution.

11K
Stars
87/100
Confianza
Categoría: ml-automationAuditoría

OpenVINO™ is an open source toolkit for optimizing and deploying AI inference

10K
Stars
82/100
Confianza
Categoría: media-automationAuditoría

Swap GPT for any LLM by changing a single line of code. Xinference lets you run open-source, speech, and multimodal models on cloud, on-prem, or your laptop — all through one unified, production-ready inference API.

9.3K
Stars
79/100
Confianza
Categoría: ml-automationAuditoría

The easiest way to serve AI apps and models - Build Model Inference APIs, Job queues, LLM apps, Multi-model pipelines, and more!

8.7K
Stars
86/100
Confianza
Categoría: developmentAuditoría

Ultra-high-performance, secure, all-in-one acceleration engine for developer resources

8.2K
Stars
86/100
Confianza
Categoría: agent-skillsAuditoría