No visual example yet
Explore the skillMMAudio
hkchengrex
[CVPR 2025] MMAudio: Taming Multimodal Joint Training for High-Quality Video-to-Audio Synthesis
OPENAGENTSKILL / DIRECTORY
Find a skill for your next task. Explore tools for Codex, Claude Code, Cursor and more.
10 Skills
Results: 10
No visual example yet
Explore the skillhkchengrex
[CVPR 2025] MMAudio: Taming Multimodal Joint Training for High-Quality Video-to-Audio Synthesis
No visual example yet
Explore the skillkeithito
A TensorFlow implementation of Google's Tacotron speech synthesis with pre-trained model (unofficial)
No visual example yet
Explore the skillmarytts
MARY TTS -- an open-source, multilingual text-to-speech synthesis system written in pure java
No visual example yet
Explore the skillsdkcarlos
A voice control - voice commands - speech recognition and speech synthesis javascript library. Create your own siri,google now or cortana with Google Chrome within your…
No visual example yet
Explore the skillOpenMOSS
MOSS‑TTS Family is an open‑source speech and sound generation model family from MOSI.AI and the OpenMOSS team. It is designed for high‑fidelity, high‑expressiveness, and…
No visual example yet
Explore the skillnateshmbhat
Offline Text To Speech synthesis for python
No visual example yet
Explore the skillsemperai
Amica is an open source interface for interactive communication with 3D characters with voice synthesis and speech recognition.
No visual example yet
Explore the skillermongroup
PyTorch implementation for SDEdit: Image Synthesis and Editing with Stochastic Differential Equations
No visual example yet
Explore the skillIBM
A `Neural = Symbolic` framework for sound and complete weighted real-value logic
No visual example yet
Explore the skillrishavbhandari6789
A modular multi-agent framework for video-based pain localisation and adaptive exercise recommendation. This prototype implements the core elements from the OpenRehabAge…