No visual example yet
Explore the skillMMAudio
hkchengrex
[CVPR 2025] MMAudio: Taming Multimodal Joint Training for High-Quality Video-to-Audio Synthesis
OPENAGENTSKILL / DIRECTORY
Find a skill for your next task. Explore tools for Codex, Claude Code, Cursor and more.
Find a skill for your next task. Preview examples where available.
1–9 / 9
Results: 9
No visual example yet
Explore the skillhkchengrex
[CVPR 2025] MMAudio: Taming Multimodal Joint Training for High-Quality Video-to-Audio Synthesis
No visual example yet
Explore the skillkeithito
A TensorFlow implementation of Google's Tacotron speech synthesis with pre-trained model (unofficial)
No visual example yet
Explore the skillmarytts
MARY TTS -- an open-source, multilingual text-to-speech synthesis system written in pure java
No visual example yet
Explore the skillsdkcarlos
A voice control - voice commands - speech recognition and speech synthesis javascript library. Create your own siri,google now or cortana with Google Chrome within your…
No visual example yet
Explore the skillnateshmbhat
Offline Text To Speech synthesis for python
No visual example yet
Explore the skillsemperai
Amica is an open source interface for interactive communication with 3D characters with voice synthesis and speech recognition.
No visual example yet
Explore the skillermongroup
PyTorch implementation for SDEdit: Image Synthesis and Editing with Stochastic Differential Equations
No visual example yet
Explore the skillzju3dv
[ICCV 2025] Diffuman4D: 4D Consistent Human View Synthesis from Sparse-View Videos with Spatio-Temporal Diffusion Models
No visual example yet
Explore the skillrishavbhandari6789
A modular multi-agent framework for video-based pain localisation and adaptive exercise recommendation. This prototype implements the core elements from the OpenRehabAge…