No visual example yet
Explore the skillDiffSinger
MoonInTheRiver
DiffSinger: Singing Voice Synthesis via Shallow Diffusion Mechanism (SVS & TTS); AAAI 2022; Official code
OPENAGENTSKILL / DIRECTORY
Find a skill for your next task. Explore tools for Codex, Claude Code, Cursor and more.
Find a skill for your next task. Preview examples where available.
1–16 / 20
Results: 20
No visual example yet
Explore the skillMoonInTheRiver
DiffSinger: Singing Voice Synthesis via Shallow Diffusion Mechanism (SVS & TTS); AAAI 2022; Official code
No visual example yet
Explore the skillhkchengrex
[CVPR 2025] MMAudio: Taming Multimodal Joint Training for High-Quality Video-to-Audio Synthesis
No visual example yet
Explore the skillkeithito
A TensorFlow implementation of Google's Tacotron speech synthesis with pre-trained model (unofficial)
No visual example yet
Explore the skillmarytts
MARY TTS -- an open-source, multilingual text-to-speech synthesis system written in pure java
No visual example yet
Explore the skilljik876
HiFi-GAN: Generative Adversarial Networks for Efficient and High Fidelity Speech Synthesis
No visual example yet
Explore the skillPyTorch implementation of convolutional neural networks-based text-to-speech synthesis models
No visual example yet
Explore the skillsdkcarlos
A voice control - voice commands - speech recognition and speech synthesis javascript library. Create your own siri,google now or cortana with Google Chrome within your…
No visual example yet
Explore the skillNVlabs
SANA: Efficient High-Resolution Image Synthesis with Linear Diffusion Transformer
No visual example yet
Explore the skillnateshmbhat
Offline Text To Speech synthesis for python
No visual example yet
Explore the skillTensorSpeech
:stuck_out_tongue_closed_eyes: TensorFlowTTS: Real-Time State-of-the-art Speech Synthesis for Tensorflow 2 (supported including English, French, Korean, Chinese, German…
No visual example yet
Explore the skillFireRedTeam
Open-source industrial-grade ASR models supporting Mandarin, Chinese dialects and English, achieving a new SOTA on public Mandarin ASR benchmarks, while also offering ou…
No visual example yet
Explore the skillcalesthio
Generate AI voiceovers, sound effects, and music using ElevenLabs APIs. Use when creating audio content for videos, podcasts, or games. Triggers include generating voice…
No visual example yet
Explore the skillEnemyx-net
A comprehensive ComfyUI integration for Microsoft's VibeVoice text-to-speech model, enabling high-quality single and multi-speaker voice synthesis directly within your C…
No visual example yet
Explore the skillmenyifang
Official implementation of "MIMO: Controllable Character Video Synthesis with Spatial Decomposed Modeling"
No visual example yet
Explore the skillsemperai
Amica is an open source interface for interactive communication with 3D characters with voice synthesis and speech recognition.
No visual example yet
Explore the skillictnlp
StreamSpeech is an “All in One” seamless model for offline and simultaneous speech recognition, speech translation and speech synthesis.