No visual example yet
Explore the skillMMAudio
hkchengrex
[CVPR 2025] MMAudio: Taming Multimodal Joint Training for High-Quality Video-to-Audio Synthesis
OPENAGENTSKILL / DIRECTORY
Find a skill for your next task. Explore tools for Codex, Claude Code, Cursor and more.
Find a skill for your next task. Preview examples where available.
1–16 / 17
Results: 17
No visual example yet
Explore the skillhkchengrex
[CVPR 2025] MMAudio: Taming Multimodal Joint Training for High-Quality Video-to-Audio Synthesis
No visual example yet
Explore the skillPluviobyte
Generate real-timestamp subtitle artifacts from final narration audio or merged video with a caption quality gate.
No visual example yet
Explore the skillmilvus-io
Dealing with all unstructured data, such as reverse image search, audio search, molecular search, video analysis, question and answer systems, NLP, etc.
No visual example yet
Explore the skillPluviobyte
Generate a controlled local narration workflow with auditions, version tracking, and subtitle-ready final audio.
No visual example yet
Explore the skillbackblaze-labs
Genblaze is an open source Python SDK for orchestrating generative AI media pipelines across video, audio, and image providers with built in provenance for every output.
No visual example yet
Explore the skillhuggingface
🤗 Diffusers: State-of-the-art diffusion models for image, video, and audio generation in PyTorch.
No visual example yet
Explore the skillEmily2040
Direct the model. Do not micro-manage the frame.
No visual example yet
Explore the skillEventual-Inc
High-performance data engine for AI and multimodal workloads. Process images, audio, video, and structured data at any scale
No visual example yet
Explore the skillwhitphx
Real-time video and audio processing on Streamlit
No visual example yet
Explore the skillYuan-ManX
Your AI Game Dev Hub. The ultimate resource hub for AI-powered game development tools. Discover cutting-edge LLMs, World Model, Agent, Code, Image, Texture, Shader, 3D M…
No visual example yet
Explore the skillThioJoe
Automatically translates the text of a video based on a subtitle file, and then uses AI voice services to create a new dubbed & translated audio track where the speech i…
No visual example yet
Explore the skillcalesthio
Video and audio processing with FFmpeg. Use for format conversion, resizing, compression, audio extraction, and preparing assets for Remotion. Triggers include convertin…
No visual example yet
Explore the skillcalesthio
AI music generation with ACE-Step 1.5 — background music, vocal tracks, covers, stem extraction for video production. Use when generating music, soundtracks, jingles, or…
No visual example yet
Explore the skillcalesthio
Transcribe audio to text using Azure AI Speech (Fast Transcription REST API). Use when converting audio/video to text, generating subtitles, or processing spoken content…
No visual example yet
Explore the skillFrancis-Rings
We present StableAvatar, the first end-to-end video diffusion transformer, which synthesizes infinite-length high-quality audio-driven avatar videos without any post-pro…
No visual example yet
Explore the skillwladradchenko
Wunjo CE: Face Swap, Lip Sync, Control Remove Objects & Text & Background, Restyling, Audio Separator, Clone Voice, Video Generation. Open Source, Local & Free.