venice-audio-speech
veniceai
Generate speech from text via POST /audio/speech, and clone a voice via POST /audio/voices. Covers TTS models (Kokoro, Qwen 3, xAI, Inworld, Chatterbox, Orpheus, ElevenL…
OPENAGENTSKILL / DIRECTORY
Find a skill for your next task. Explore tools for Codex, Claude Code, Cursor and more.
Find a skill for your next task. Preview examples where available.
Page 32 · 16 shown · 743 public entries
Results: 743
veniceai
Generate speech from text via POST /audio/speech, and clone a voice via POST /audio/voices. Covers TTS models (Kokoro, Qwen 3, xAI, Inworld, Chatterbox, Orpheus, ElevenL…
veniceai
Route a prompt to the right Venice text model based on privacy tier (anonymized / private / TEE / E2EE), modality (vision / audio / video input), capability (reasoning,…
DemisEom
A Implementation of SpecAugment with Tensorflow & Pytorch, introduced by Google Brain
eduardolat
🔊 Kokoro Web: Free AI text-to-speech, online or self-hosted, OpenAI compatible!
heygen-com
Translate and dub a video into another language with voice cloning and lip-sync, powered by HeyGen Video Translation. The presenter keeps their face, their voice is clon…
daniilrobnikov
VITS2: Improving Quality and Efficiency of Single-Stage Text-to-Speech with Adversarial Learning and Architecture Design
evancohen
:speech_balloon: /so.nus/ STT (speech to text) for Node with offline hotword detection
jaccen
MCP protocol integration with 3DGS rendering pipeline: Agent-controlled Three.js/WebGPU rendering, voice-driven scene reconstruction, real-time parameter manipulation, l…
sooftware
Open-Source Toolkit for End-to-End Korean Automatic Speech Recognition leveraging PyTorch and Hydra.
lucasnewman
Implementation of F5-TTS in MLX
zenstory-ai
Direct game art and creative vision. Turn GAME_DESIGN into ART_DIRECTION defining a recognizable visual identity, functional visual and audio feedback, and key in-game m…
alyssaxuu
Supercharge Notion with custom commands to record, draw, and more ✍️
reriiasu
Real-time transcription using faster-whisper
jianshuo
Use when the user has a video + an SRT and wants the subtitles either burned into the pixels (libass, always-visible) or soft-muxed as a togglable track. Also handles th…
naive-kun
End-to-end, stateful talking-head video production for complete beginners and experienced editors. Use when a user wants optional rough-cut guidance, transcription or ca…
huawei-noah
This is the main repository of open-sourced speech technology by Huawei Noah's Ark Lab.