No visual example yet
Explore the skillfish-audio-tts
calesthio
Generate expressive, multilingual narration with fish.audio (S1 / S2-generation models) and reuse cloned voices via reference_id. Use when the user prefers fish.audio/Fi…
OPENAGENTSKILL / DIRECTORY
Find a skill for your next task. Explore tools for Codex, Claude Code, Cursor and more.
70 Skills
Results: 70
No visual example yet
Explore the skillcalesthio
Generate expressive, multilingual narration with fish.audio (S1 / S2-generation models) and reuse cloned voices via reference_id. Use when the user prefers fish.audio/Fi…
No visual example yet
Explore the skillPluviobyte
Generate real-timestamp subtitle artifacts from final narration audio or merged video with a caption quality gate.
No visual example yet
Explore the skillhuggingface
🤗 Transformers: the model-definition framework for state-of-the-art machine learning models in text, vision, audio, and multimodal models, for both inference and traini…
No visual example yet
Explore the skillHumanSignal
Label Studio is a multi-type data labeling and annotation tool with standardized output format
No visual example yet
Explore the skillabus-aikorea
Gradio WebUI for creators and developers, featuring key TTS (Edge-TTS, kokoro) and zero-shot Voice Cloning (E2 & F5-TTS, CosyVoice), with Whisper audio processing, YouTu…
No visual example yet
Explore the skillopen-mmlab
Amphion (/æmˈfaɪən/) is a toolkit for Audio, Music, and Speech Generation. Its purpose is to support reproducible research and help junior researchers and engineers get…
No visual example yet
Explore the skillspotify
No visual example yet
Explore the skillspotify
A lightweight yet powerful audio-to-MIDI converter with pitch bend detection
No visual example yet
Explore the skillrsxdalv
A single Gradio + React WebUI with extensions for ACE-Step, OmniVoice, Kimi Audio, Piper TTS, GPT-SoVITS, CosyVoice, XTTSv2, DIA, Kokoro, OpenVoice, ParlerTTS, Stable Au…
No visual example yet
Explore the skillmilvus-io
Dealing with all unstructured data, such as reverse image search, audio search, molecular search, video analysis, question and answer systems, NLP, etc.
No visual example yet
Explore the skillxtreme1-io
Xtreme1 is an all-in-one data labeling and annotation platform for multimodal data training and supports 3D LiDAR point cloud, image, and LLM.
No visual example yet
Explore the skillCPJKU
Python audio and music signal processing library
No visual example yet
Explore the skillpluja
Transcribe any audio to text, translate and edit subtitles 100% locally with a web UI. Powered by whisper models!
No visual example yet
Explore the skillyatengLG
Labeling tool with SAM(segment anything model),supports SAM, SAM2, SAM3, sam-hq, MobileSAM EdgeSAM etc.交互式半自动图像标注工具
No visual example yet
Explore the skillenhuiz
An unofficial PyTorch implementation of the audio LM VALL-E
No visual example yet
Explore the skillreadbeyond
aeneas is a Python/C library and a set of tools to automagically synchronize audio and text (aka forced alignment)