No visual example yet
Explore the skillMatcha TTS
shivammehta25
[ICASSP 2024] 🍵 Matcha-TTS: A fast TTS architecture with conditional flow matching
OPENAGENTSKILL / DIRECTORY
Find a skill for your next task. Explore tools for Codex, Claude Code, Cursor and more.
Find a skill for your next task. Preview examples where available.
145–160 / 466
Results: 466
No visual example yet
Explore the skillshivammehta25
[ICASSP 2024] 🍵 Matcha-TTS: A fast TTS architecture with conditional flow matching
No visual example yet
Explore the skillFunAudioLLM
End-to-end speech recognition large model: 31 languages, dialects, accents, lyrics, hotwords, timestamps, speaker diarization. Trained on tens of millions of hours.
No visual example yet
Explore the skilldevnen
Self-host the powerful Chatterbox TTS model. This server offers a user-friendly Web UI, flexible API endpoints (incl. OpenAI compatible), predefined voices, voice clonin…
No visual example yet
Explore the skillRenumics
Interactively explore unstructured datasets from your dataframe.
No visual example yet
Explore the skillYuan-ManX
Your AI Game Dev Hub. The ultimate resource hub for AI-powered game development tools. Discover cutting-edge LLMs, World Model, Agent, Code, Image, Texture, Shader, 3D M…
No visual example yet
Explore the skillmodal-labs
No visual example yet
Explore the skillmyshell-ai
Instant voice cloning by MIT and MyShell. Audio foundation model.
No visual example yet
Explore the skillphuc-nt
Real-time speech translation — macOS & Windows, free TTS, no server, your API keys only
No visual example yet
Explore the skilljanvarev
Ирина - русский голосовой ассистент для работы оффлайн. Поддерживает скиллы через плагины.
No visual example yet
Explore the skilllhotse-speech
Tools for handling multimodal data in machine learning projects.
No visual example yet
Explore the skilletro-js
Typescript video-editing library for the browser
No visual example yet
Explore the skilldiodiogod
A ComfyUI custom node integration for local multi-engine multi-language Text-to-Speech and Voice Conversion. Supports: RVC, Echo-TTS, Qwen3-TTS, Cozy Voice 3, Step Audio…
No visual example yet
Explore the skilllibAudioFlux
A library for audio and music analysis, feature extraction.
No visual example yet
Explore the skillrsxdalv
A single Gradio + React WebUI with extensions for ACE-Step, OmniVoice, Kimi Audio, Piper TTS, GPT-SoVITS, CosyVoice, XTTSv2, DIA, Kokoro, OpenVoice, ParlerTTS, Stable Au…
No visual example yet
Explore the skillJimmySadek
Claude Code skill: turn YouTube videos into structured, Obsidian-ready Markdown notes with full metadata, chapters, and transcripts
No visual example yet
Explore the skillstemrollerapp
Isolate vocals, drums, bass, and other instrumental stems from any song
Design direction, Figma implementation, UI review, React performance, browser QA, and safe preview deployments.
Code review, repo inspection, testing, planning, shipping, and engineering workflows for coding agents.
Research, recent-events briefings, PDF parsing, markdown conversion, RAG ingestion, and knowledge workflows.
Stock analysis, market research, quant backtesting, financial data, and investment research skills.
Crawling, scraping, extraction, browser automation, structured data capture, and website-to-markdown workflows.
Presentation generation, editable PPTX, slide decks, speaker notes, and visual storytelling workflows.
Choose React motion graphics, generated footage or editing; compare dependencies and inspect actual outputs.
Image, video, creative production, UI design, multimodal generation, and visual workflow skills.