GPT SoVITS
RVC-Boss
1 min voice data can also be used to train a good TTS model! (few shot voice cloning)
OPENAGENTSKILL / DIRECTORY
Find a skill for your next task. Explore tools for Codex, Claude Code, Cursor and more.
33–48 / 87
Candidates in this shortlist, not the full registry. GitHub stars belong to repositories, not individual skills.
Results: 87
RVC-Boss
1 min voice data can also be used to train a good TTS model! (few shot voice cloning)
babysor
🚀Clone a voice in 5 seconds to generate arbitrary speech in real-time
metavoiceio
Foundational model for human-like, expressive TTS
lenML
🍦 Speech-AI-Forge is a project developed around TTS generation model, implementing an API Server and a Gradio-based WebUI.
Finrandojin
AI-powered multi-voice audiobook generator — LLM script annotation, voice cloning, voice design, LoRA training, per-line style control, and export to MP3, chaptered M4B,…
runesleo
Agent Skill + Remotion pipeline: brief/script → review receipt → narrated 9:16 explainer. RC: video-explainer skill.
BinWang28
The hub for audio AI research: papers, open models, benchmarks & datasets across audio LLMs, speech recognition, TTS, music & audio generation.
AutoArk
[AutoArk] GPA (General Purpose Audio) can do ASR, TTS and voice conversion with one tiny model!
nari-labs
A TTS model capable of generating ultra-realistic dialogue in one pass.
richardr1126
An open-source read-along document reader server with high-quality TTS options, synchronized highlighting, and audiobook export for EPUB, PDF, DOCX, TXT, and MD.
gexgd0419
Make Azure natural TTS voices accessible to any SAPI 5-compatible application.
StarmoonAI
A conversational, AI device + software framework for companionship, entertainment, education, healthcare, IoT applications, and DIY robotics. Built with Python, NextJS,…
herimor
VoXtream is a Full-Stream Zero-shot TTS model with Extremely Low Latency and Speaking rate Control