OPENAGENTSKILL / DIRECTORY
Find a skill for your next task. Explore tools for Codex, Claude Code, Cursor and more.
81–96 / 164
Candidates in this shortlist, not the full registry. GitHub stars belong to repositories, not individual skills.
Results: 164
pannous
🎙Speech recognition using the tensorflow deep learning framework, sequence-to-sequence neural networks
Enemyx-net
A comprehensive ComfyUI integration for Microsoft's VibeVoice text-to-speech model, enabling high-quality single and multi-speaker voice synthesis directly within your C…
Phantom-video
Phantom: Subject-Consistent Video Generation via Cross-Modal Alignment
Azure-Samples
Microsoft Text-to-Speech API sample code in several languages, part of Cognitive Services.
travisvn
Free, high-quality text-to-speech API endpoint to replace OpenAI, Azure, or ElevenLabs
bytedance
[ICCV 2025] 🔥🔥 UNO: A Universal Customization Method for Both Single and Multi-Subject Conditioning
Softcatala
Whisper command line client compatible with original OpenAI client based on CTranslate2.
ekwek1
Soprano: Instant, Ultra-Realistic Text-to-Speech
kan-bayashi
Unofficial Parallel WaveGAN (+ MelGAN & Multi-band MelGAN & HiFi-GAN & StyleMelGAN) with Pytorch
nari-labs
TTS model capable of streaming conversational audio in realtime.
Saik0s
The open-source iOS app that's making quality voice transcription more accessible on mobile devices.
microsoft
Unified-Modal Speech-Text Pre-Training for Spoken Language Processing
voice-cloning-app
A Python/Pytorch app for easily synthesising human voices
coqui-ai
💎 A list of accessible speech corpora for ASR, TTS, and other Speech Technologies