No visual example yet
Explore the skillazure-speech-to-text
calesthio
Transcribe audio to text using Azure AI Speech (Fast Transcription REST API). Use when converting audio/video to text, generating subtitles, or processing spoken content…
OPENAGENTSKILL / DIRECTORY
Find a skill for your next task. Explore tools for Codex, Claude Code, Cursor and more.
Find a skill for your next task. Preview examples where available.
49–64 / 64
Results: 64
No visual example yet
Explore the skillcalesthio
Transcribe audio to text using Azure AI Speech (Fast Transcription REST API). Use when converting audio/video to text, generating subtitles, or processing spoken content…
No visual example yet
Explore the skillcalesthio
Generate neural narration audio using Azure AI Speech (REST text-to-speech). Use when synthesizing voiceovers or narration in OpenMontage. Optional cloud TTS provider —…
No visual example yet
Explore the skillcalesthio
Generate AI voiceovers, sound effects, and music using ElevenLabs APIs. Use when creating audio content for videos, podcasts, or games. Triggers include generating voice…
No visual example yet
Explore the skillastorfi
:unlock: Lip Reading - Cross Audio-Visual Recognition using 3D Architectures
No visual example yet
Explore the skillFrancis-Rings
We present StableAvatar, the first end-to-end video diffusion transformer, which synthesizes infinite-length high-quality audio-driven avatar videos without any post-pro…
No visual example yet
Explore the skillwladradchenko
Wunjo CE: Face Swap, Lip Sync, Control Remove Objects & Text & Background, Restyling, Audio Separator, Clone Voice, Video Generation. Open Source, Local & Free.
No visual example yet
Explore the skillnari-labs
TTS model capable of streaming conversational audio in realtime.
No visual example yet
Explore the skillmikeal
📞 Free and reliable audio calls for everyone w/ browser p2p.
No visual example yet
Explore the skillMengTo
Generate ElevenLabs text-to-speech audio from scripts or inline text using local voice profiles. Use when the user asks for ElevenLabs, text-to-speech, TTS, narration, v…
No visual example yet
Explore the skillAutoArk
[AutoArk] GPA (General Purpose Audio) can do ASR, TTS and voice conversion with one tiny model!
No visual example yet
Explore the skillelementalsouls
Hunt CAPTCHA Bypass — 6 distinct patterns: (1) CAPTCHA field simply omitted from the request (server-side validation absent), (2) CAPTCHA token replayed from a solved ch…
No visual example yet
Explore the skillalesaccoia
Near-Realtime audio transcription using self-hosted Whisper and WebSocket in Python/JS
No visual example yet
Explore the skillpexoai
A collection of open-source Agent Skills for content creation — images, audio, and video.
No visual example yet
Explore the skillpexoai
Expert prompt engineering for Seedance 2.0. Use when the user wants to generate a video with multimodal assets (images, videos, audio) and needs the best possible prompt.
No visual example yet
Explore the skillhomelab-00
A fully local and private Speech-To-Text app with cross-platform support, speaker diarization, Audio Notebook mode, LM Studio integration, and both longform and live tra…
No visual example yet
Explore the skillSYuan03
Any source (PDF, video, web, audio, text) to interactive learning package with quizzes, flashcards and spaced repetition. One command, 12-section study guide.