No visual example yet
Explore the skillDiffSinger
MoonInTheRiver
DiffSinger: Singing Voice Synthesis via Shallow Diffusion Mechanism (SVS & TTS); AAAI 2022; Official code
OPENAGENTSKILL / DIRECTORY
Find a skill for your next task. Explore tools for Codex, Claude Code, Cursor and more.
29 Skills
Results: 29
No visual example yet
Explore the skillMoonInTheRiver
DiffSinger: Singing Voice Synthesis via Shallow Diffusion Mechanism (SVS & TTS); AAAI 2022; Official code
No visual example yet
Explore the skillhkchengrex
[CVPR 2025] MMAudio: Taming Multimodal Joint Training for High-Quality Video-to-Audio Synthesis
No visual example yet
Explore the skillkeithito
A TensorFlow implementation of Google's Tacotron speech synthesis with pre-trained model (unofficial)
No visual example yet
Explore the skillmarytts
MARY TTS -- an open-source, multilingual text-to-speech synthesis system written in pure java
No visual example yet
Explore the skilljik876
HiFi-GAN: Generative Adversarial Networks for Efficient and High Fidelity Speech Synthesis
No visual example yet
Explore the skillPyTorch implementation of convolutional neural networks-based text-to-speech synthesis models
No visual example yet
Explore the skillsdkcarlos
A voice control - voice commands - speech recognition and speech synthesis javascript library. Create your own siri,google now or cortana with Google Chrome within your…
No visual example yet
Explore the skillaashaexo
Combines pattern detection, statistical writing signals, and voice calibration to make drafts sound more human.
No visual example yet
Explore the skillcalesthio
Generate AI voiceovers, sound effects, and music using ElevenLabs APIs. Use when creating audio content for videos, podcasts, or games. Triggers include generating voice…
No visual example yet
Explore the skillNVlabs
SANA: Efficient High-Resolution Image Synthesis with Linear Diffusion Transformer
No visual example yet
Explore the skillOpenMOSS
MOSS‑TTS Family is an open‑source speech and sound generation model family from MOSI.AI and the OpenMOSS team. It is designed for high‑fidelity, high‑expressiveness, and…
No visual example yet
Explore the skillnateshmbhat
Offline Text To Speech synthesis for python
No visual example yet
Explore the skillTensorSpeech
:stuck_out_tongue_closed_eyes: TensorFlowTTS: Real-Time State-of-the-art Speech Synthesis for Tensorflow 2 (supported including English, French, Korean, Chinese, German…
No visual example yet
Explore the skillnexu-io
Audio generation skill — jingles, beds, voiceover, and sound effects. Routes music requests to Suno V5 / Udio / Lyria, speech to MiniMax TTS / FishAudio / ElevenLabs V3,…

FoundationVision
[CVPR 2025 Oral]Infinity ∞ : Scaling Bitwise AutoRegressive Modeling for High-Resolution Image Synthesis
View previews · 3No visual example yet
Explore the skillK-Dense-AI
Search scientific papers and retrieve structured experimental data extracted from full-text studies via the BGPT MCP server. Returns 25+ fields per paper including metho…