Confucius4 TTS
netease-youdao
Confucius4-TTS: a Multilingual and Cross-Lingual Zero-Shot TTS Engine
OPENAGENTSKILL / DIRECTORY
Find a skill for your next task. Explore tools for Codex, Claude Code, Cursor and more.
49–58 / 58
Candidates in this shortlist, not the full registry. GitHub stars belong to repositories, not individual skills.
Results: 58
netease-youdao
Confucius4-TTS: a Multilingual and Cross-Lingual Zero-Shot TTS Engine
huggingface
🤗 Diffusers: State-of-the-art diffusion models for image, video, and audio generation in PyTorch.
myshell-ai
Instant voice cloning by MIT and MyShell. Audio foundation model.
pnnbao97
Vietnamese TTS with instant voice cloning • On-device • Real-time CPU inference • 24kHz audio quality • Chuyển văn bản thành giọng nói tiếng Việt • Text to speech tiếng…
ThioJoe
Automatically translates the text of a video based on a subtitle file, and then uses AI voice services to create a new dubbed & translated audio track where the speech i…
Francis-Rings
We present StableAvatar, the first end-to-end video diffusion transformer, which synthesizes infinite-length high-quality audio-driven avatar videos without any post-pro…
wladradchenko
Wunjo CE: Face Swap, Lip Sync, Control Remove Objects & Text & Background, Restyling, Audio Separator, Clone Voice, Video Generation. Open Source, Local & Free.
nari-labs
TTS model capable of streaming conversational audio in realtime.
alesaccoia
Near-Realtime audio transcription using self-hosted Whisper and WebSocket in Python/JS
AutoArk
[AutoArk] GPA (General Purpose Audio) can do ASR, TTS and voice conversion with one tiny model!