Dia TTS Server
devnen
Self-host the powerful Dia TTS model. This server offers a user-friendly Web UI, flexible API endpoints (incl. OpenAI compatible), support for SafeTensors/BF16, voice cl…
OPENAGENTSKILL / DIRECTORY
Find a skill for your next task. Explore tools for Codex, Claude Code, Cursor and more.
1–16 / 87
Candidates in this shortlist, not the full registry. GitHub stars belong to repositories, not individual skills.
Results: 87
devnen
Self-host the powerful Dia TTS model. This server offers a user-friendly Web UI, flexible API endpoints (incl. OpenAI compatible), support for SafeTensors/BF16, voice cl…
index-tts
An Industrial-Level Controllable and Efficient Zero-Shot Text-To-Speech System
coqui-ai
🐸💬 - a deep learning toolkit for Text-to-Speech, battle-tested in research and production
rany2
Use Microsoft Edge's online text-to-speech service from Python WITHOUT needing Microsoft Edge or Windows or an API key
mozilla
:robot: :speech_balloon: Deep learning for Text to Speech (Discussion forum: https://discourse.mozilla.org/c/tts)
pnnbao97
Vietnamese TTS with instant voice cloning • On-device • Real-time CPU inference • 24kHz audio quality • Chuyển văn bản thành giọng nói tiếng Việt • Text to speech tiếng…
devnen
Self-host the powerful Chatterbox TTS model. This server offers a user-friendly Web UI, flexible API endpoints (incl. OpenAI compatible), predefined voices, voice clonin…
nazdridoy
A CLI text-to-speech tool using the Kokoro model, supporting multiple languages, voices (with blending), and various input formats including EPUB books and PDF documents.
marytts
MARY TTS -- an open-source, multilingual text-to-speech synthesis system written in pure java
Azure-Samples
Microsoft Text-to-Speech API sample code in several languages, part of Cognitive Services.
travisvn
Free, high-quality text-to-speech API endpoint to replace OpenAI, Azure, or ElevenLabs
Saganaki22
OmniVoice TTS nodes for ComfyUI - Zero-shot multilingual text-to-speech with voice cloning, voice design, and multi-speaker dialogue
Pluviobyte
Generate a controlled local narration workflow with auditions, version tracking, and subtitle-ready final audio.
Aratako
A Flow Matching-based Text-to-Speech Model with Emoji-driven Style Control