Rhino
Picovoice
On-device Speech-to-Intent engine powered by deep learning
OPENAGENTSKILL / DIRECTORY
Find a skill for your next task. Explore tools for Codex, Claude Code, Cursor and more.
1–16 / 40
Candidates in this shortlist, not the full registry. GitHub stars belong to repositories, not individual skills.
Results: 40
Picovoice
On-device Speech-to-Intent engine powered by deep learning
FunAudioLLM
Multilingual speech understanding: ASR + emotion recognition + audio event detection. 50+ languages, 15x faster than Whisper, non-autoregressive.
pnnbao97
Vietnamese TTS with instant voice cloning • On-device • Real-time CPU inference • 24kHz audio quality • Chuyển văn bản thành giọng nói tiếng Việt • Text to speech tiếng…
thu-ml
[ICLR2025, ICML2025, NeurIPS2025 Spotlight] Quantized Attention achieves speedup of 2-5x compared to FlashAttention, without losing end-to-end metrics across language, i…
argmaxinc
On-device Speech AI for Apple Silicon
NVIDIA
State-of-the-Art Deep Learning scripts organized by models - easy to train and deploy with reproducible accuracy and performance on enterprise-grade infrastructure.
vllm-project
A framework for efficient model inference with omni-modality models
hao-ai-lab
A unified inference and post-training framework for accelerated video generation.
thu-ml
TurboDiffusion: 100–200× Acceleration for Video Diffusion Models
bytedance
SALMONN family: A suite of advanced multi-modal LLMs
ARahim3
Fine-tune LLMs on your Mac with Apple Silicon. SFT, DPO, GRPO, Vision, TTS, STT, Embedding, and OCR fine-tuning — natively on MLX. Unsloth-compatible API.
FunAudioLLM
End-to-end speech recognition large model: 31 languages, dialects, accents, lyrics, hotwords, timestamps, speaker diarization. Trained on tens of millions of hours.