No visual example yet
Explore the skillTTS Audio Suite
diodiogod
A ComfyUI custom node integration for local multi-engine multi-language Text-to-Speech and Voice Conversion. Supports: RVC, Echo-TTS, Qwen3-TTS, Cozy Voice 3, Step Audio…
OPENAGENTSKILL / DIRECTORY
Trouvez un skill pour votre prochaine tâche avec Codex, Claude Code, Cursor et plus encore.
104 Skills
Résultats: 104
No visual example yet
Explore the skilldiodiogod
A ComfyUI custom node integration for local multi-engine multi-language Text-to-Speech and Voice Conversion. Supports: RVC, Echo-TTS, Qwen3-TTS, Cozy Voice 3, Step Audio…
No visual example yet
Explore the skillFunAudioLLM
Multi-lingual large voice generation model, providing inference, training and deployment full-stack ability.
No visual example yet
Explore the skillopen-mmlab
Amphion (/æmˈfaɪən/) is a toolkit for Audio, Music, and Speech Generation. Its purpose is to support reproducible research and help junior researchers and engineers get…
No visual example yet
Explore the skillrsxdalv
A single Gradio + React WebUI with extensions for ACE-Step, OmniVoice, Kimi Audio, Piper TTS, GPT-SoVITS, CosyVoice, XTTSv2, DIA, Kokoro, OpenVoice, ParlerTTS, Stable Au…
No visual example yet
Explore the skillBlaizzy
A text-to-speech (TTS), speech-to-text (STT) and speech-to-speech (STS) library built on Apple's MLX framework, providing efficient speech analysis on Apple Silicon.
No visual example yet
Explore the skillpytorch
Data manipulation and transformation for audio signal processing, powered by PyTorch
No visual example yet
Explore the skilllibAudioFlux
A library for audio and music analysis, feature extraction.

yanliudesign
Generate original one-ink or controlled two-ink editorial images from any theme, sentence, article idea, object, or reference photo. Always use this skill when the user…
Voir les exemples · 1No visual example yet
Explore the skillhuggingface
🤗 Diffusers: State-of-the-art diffusion models for image, video, and audio generation in PyTorch.
No visual example yet
Explore the skillFunAudioLLM
Multilingual speech understanding: ASR + emotion recognition + audio event detection. 50+ languages, 15x faster than Whisper, non-autoregressive.
No visual example yet
Explore the skillTencent-Hunyuan
HunyuanVideo: A Systematic Framework For Large Video Generation Model
No visual example yet
Explore the skillPluviobyte
Generate real-timestamp subtitle artifacts from final narration audio or merged video with a caption quality gate.
Alisa0808
Turn one topic into a narrated Vox-style paper-collage explainer or ad video, from script through captions.
Voir les exemples · 3No visual example yet
Explore the skillHKUDS
"ViMax: Agentic Video Generation (Director, Screenwriter, Producer, and Video Generator All-in-One)"
No visual example yet
Explore the skillduixcom
🚀 Truly open-source AI avatar(digital human) toolkit for offline video generation and digital human cloning.
No visual example yet
Explore the skillVectorSpaceLab
OmniGen: Unified Image Generation. https://arxiv.org/pdf/2409.11340