GPT SoVITS
RVC-Boss
1 min voice data can also be used to train a good TTS model! (few shot voice cloning)
OPENAGENTSKILL / DIRECTORY
Find a skill for your next task. Explore tools for Codex, Claude Code, Cursor and more.
1–16 / 319
Candidates in this shortlist, not the full registry. GitHub stars belong to repositories, not individual skills.
Results: 319
RVC-Boss
1 min voice data can also be used to train a good TTS model! (few shot voice cloning)
babysor
🚀Clone a voice in 5 seconds to generate arbitrary speech in real-time
AIDC-AI
🚀 AI 全自动短视频引擎 | AI Fully Automated Short Video Engine
m-bain
WhisperX: Automatic Speech Recognition with Word-level Timestamps (& Diarization)
jianchang512
Translate the video from one language to another and embed dubbing & subtitles.
Wan-Video
Wan: Open and Advanced Large-Scale Video Generative Models
HKUDS
"ViMax: Agentic Video Generation (Director, Screenwriter, Producer, and Video Generator All-in-One)"
nari-labs
A TTS model capable of generating ultra-realistic dialogue in one pass.
myshell-ai
Instant voice cloning by MIT and MyShell. Audio foundation model.
Tencent-Hunyuan
HunyuanVideo: A Systematic Framework For Large Video Generation Model
Uberi
Speech recognition module for Python, supporting several engines and APIs, online and offline.
rany2
Use Microsoft Edge's online text-to-speech service from Python WITHOUT needing Microsoft Edge or Windows or an API key