E2 Tts Pytorch
lucidrains
Implementation of E2-TTS, "Embarrassingly Easy Fully Non-Autoregressive Zero-Shot TTS", in Pytorch
OPENAGENTSKILL / DIRECTORY
Find a skill for your next task. Explore tools for Codex, Claude Code, Cursor and more.
Find a skill for your next task. Preview examples where available.
33–48 / 78
Results: 78
lucidrains
Implementation of E2-TTS, "Embarrassingly Easy Fully Non-Autoregressive Zero-Shot TTS", in Pytorch
devnen
Self-host the ultra-lightweight Kitten TTS model with this enhanced API server with an intuitive Web UI, large text processing for audiobooks, and GPU acceleration.
Pluviobyte
Generate a controlled local narration workflow with auditions, version tracking, and subtitle-ready final audio.
petermg
Modified version of Chatterbox that accepts text files as input and no character restrictions. I use it to make audiobooks, especially for my kids.
RVC-Boss
1 min voice data can also be used to train a good TTS model! (few shot voice cloning)
metavoiceio
Foundational model for human-like, expressive TTS
lenML
🍦 Speech-AI-Forge is a project developed around TTS generation model, implementing an API Server and a Gradio-based WebUI.
fatchord
AutoArk
[AutoArk] GPA (General Purpose Audio) can do ASR, TTS and voice conversion with one tiny model!
BinWang28
The hub for audio AI research: papers, open models, benchmarks & datasets across audio LLMs, speech recognition, TTS, music & audio generation.
gexgd0419
Make Azure natural TTS voices accessible to any SAPI 5-compatible application.
kxxt
A simple text-to-speech client for Azure TTS API.
wildminder
ComfyUI custom node for the VibeVoice TTS. Expressive, long-form, multi-speaker conversational audio
shibing624
Automatic Speech Recognition(ASR), Text-To-Speech(TTS) engine. 中英语音识别、多角色语音合成,支持多语言,准确率高
herimor
VoXtream is a Full-Stream Zero-shot TTS model with Extremely Low Latency and Speaking rate Control
frostming
A unified interface for multiple Text-to-Speech (TTS) providers.