No visual example yet
Explore the skillacestep
calesthio
AI music generation with ACE-Step 1.5 — background music, vocal tracks, covers, stem extraction for video production. Use when generating music, soundtracks, jingles, or…
OPENAGENTSKILL / DIRECTORY
Find a skill for your next task. Explore tools for Codex, Claude Code, Cursor and more.
69 Skills
Results: 69
No visual example yet
Explore the skillcalesthio
AI music generation with ACE-Step 1.5 — background music, vocal tracks, covers, stem extraction for video production. Use when generating music, soundtracks, jingles, or…
No visual example yet
Explore the skillcalesthio
Transcribe audio to text using Azure AI Speech (Fast Transcription REST API). Use when converting audio/video to text, generating subtitles, or processing spoken content…
No visual example yet
Explore the skillcalesthio
Generate neural narration audio using Azure AI Speech (REST text-to-speech). Use when synthesizing voiceovers or narration in OpenMontage. Optional cloud TTS provider —…
No visual example yet
Explore the skillcalesthio
Generate AI voiceovers, sound effects, and music using ElevenLabs APIs. Use when creating audio content for videos, podcasts, or games. Triggers include generating voice…
No visual example yet
Explore the skillKichangKim
AI based multi-label girl image classification system, implemented by using TensorFlow.
No visual example yet
Explore the skillBrikerMan
Kashgari is a production-level NLP Transfer learning framework built on top of tf.keras for text-labeling and text-classification, includes Word2Vec, BERT, and GPT2 Lang…
No visual example yet
Explore the skillastorfi
:unlock: Lip Reading - Cross Audio-Visual Recognition using 3D Architectures
No visual example yet
Explore the skillszilard
A minimal benchmark for scalability, speed and accuracy of commonly used open source implementations (R packages, Python scikit-learn, H2O, xgboost, Spark MLlib etc.) of…
No visual example yet
Explore the skillFrancis-Rings
We present StableAvatar, the first end-to-end video diffusion transformer, which synthesizes infinite-length high-quality audio-driven avatar videos without any post-pro…
No visual example yet
Explore the skillwladradchenko
Wunjo CE: Face Swap, Lip Sync, Control Remove Objects & Text & Background, Restyling, Audio Separator, Clone Voice, Video Generation. Open Source, Local & Free.
No visual example yet
Explore the skillnari-labs
TTS model capable of streaming conversational audio in realtime.
No visual example yet
Explore the skillmikeal
📞 Free and reliable audio calls for everyone w/ browser p2p.
No visual example yet
Explore the skillMengTo
Generate ElevenLabs text-to-speech audio from scripts or inline text using local voice profiles. Use when the user asks for ElevenLabs, text-to-speech, TTS, narration, v…
No visual example yet
Explore the skillAutoArk
[AutoArk] GPA (General Purpose Audio) can do ASR, TTS and voice conversion with one tiny model!
No visual example yet
Explore the skillalesaccoia
Near-Realtime audio transcription using self-hosted Whisper and WebSocket in Python/JS
No visual example yet
Explore the skillpexoai
A collection of open-source Agent Skills for content creation — images, audio, and video.