OPENAGENTSKILL / DIRECTORY
Find a skill for your next task. Explore tools for Codex, Claude Code, Cursor and more.
193–208 / 478
Candidates in this shortlist, not the full registry. GitHub stars belong to repositories, not individual skills.
Results: 478
pannous
🎙Speech recognition using the tensorflow deep learning framework, sequence-to-sequence neural networks
MetalPetal
A GPU accelerated image and video processing framework built on Metal.
Enemyx-net
A comprehensive ComfyUI integration for Microsoft's VibeVoice text-to-speech model, enabling high-quality single and multi-speaker voice synthesis directly within your C…
Phantom-video
Phantom: Subject-Consistent Video Generation via Cross-Modal Alignment
Azure-Samples
Microsoft Text-to-Speech API sample code in several languages, part of Cognitive Services.
thu-ml
[ICML2025] SpargeAttention: A training-free sparse attention that accelerates any model inference.
superstreamerapp
An open, scalable, online streaming setup. All-in-one toolkit from ingest to adaptive video playback. Built for developers in need of video tooling.
FoundationVision
Autoregressive Model Beats Diffusion: 🦙 Llama for Scalable Image Generation
travisvn
Free, high-quality text-to-speech API endpoint to replace OpenAI, Azure, or ElevenLabs
julius-speech
Open-Source Large Vocabulary Continuous Speech Recognition Engine
bytedance
[ICCV 2025] 🔥🔥 UNO: A Universal Customization Method for Both Single and Multi-Subject Conditioning
Softcatala
Whisper command line client compatible with original OpenAI client based on CTranslate2.
syhw
Attempt at tracking states of the arts and recent results (bibliography) on speech recognition.