No visual example yet
Explore the skillVoice AI
rapidaai
Rapida is an open-source, end-to-end voice AI orchestration platform for building real-time conversational voice agents with audio streaming, STT, TTS, VAD, multi-channe…
OPENAGENTSKILL / DIRECTORY
Find a skill for your next task. Explore tools for Codex, Claude Code, Cursor and more.
Find a skill for your next task. Preview examples where available.
33–48 / 55
Results: 55
No visual example yet
Explore the skillrapidaai
Rapida is an open-source, end-to-end voice AI orchestration platform for building real-time conversational voice agents with audio streaming, STT, TTS, VAD, multi-channe…
No visual example yet
Explore the skillinfiniV
Local voice dictation and meeting recorder for Windows + Linux. Hold a hotkey to dictate, or record long-form meetings with system audio. Whisper transcription, bring-yo…
No visual example yet
Explore the skillhuggingface
🤗 Diffusers: State-of-the-art diffusion models for image, video, and audio generation in PyTorch.
No visual example yet
Explore the skillPixVerseAI
Agent skill library for PixVerse CLI — helps AI agents (Claude Code, Cursor, Codex, etc.) generate videos and images through structured, composable workflows.
No visual example yet
Explore the skillEmily2040
Direct the model. Do not micro-manage the frame.
No visual example yet
Explore the skillmyshell-ai
Instant voice cloning by MIT and MyShell. Audio foundation model.
No visual example yet
Explore the skillwhitphx
Real-time video and audio processing on Streamlit
No visual example yet
Explore the skillpnnbao97
Vietnamese TTS with instant voice cloning • On-device • Real-time CPU inference • 24kHz audio quality • Chuyển văn bản thành giọng nói tiếng Việt • Text to speech tiếng…
No visual example yet
Explore the skillThioJoe
Automatically translates the text of a video based on a subtitle file, and then uses AI voice services to create a new dubbed & translated audio track where the speech i…
No visual example yet
Explore the skillcalesthio
Video and audio processing with FFmpeg. Use for format conversion, resizing, compression, audio extraction, and preparing assets for Remotion. Triggers include convertin…
No visual example yet
Explore the skillcalesthio
AI music generation with ACE-Step 1.5 — background music, vocal tracks, covers, stem extraction for video production. Use when generating music, soundtracks, jingles, or…
No visual example yet
Explore the skillcalesthio
Transcribe audio to text using Azure AI Speech (Fast Transcription REST API). Use when converting audio/video to text, generating subtitles, or processing spoken content…
No visual example yet
Explore the skillcalesthio
Generate neural narration audio using Azure AI Speech (REST text-to-speech). Use when synthesizing voiceovers or narration in OpenMontage. Optional cloud TTS provider —…
No visual example yet
Explore the skillcalesthio
Generate AI voiceovers, sound effects, and music using ElevenLabs APIs. Use when creating audio content for videos, podcasts, or games. Triggers include generating voice…
No visual example yet
Explore the skillastorfi
:unlock: Lip Reading - Cross Audio-Visual Recognition using 3D Architectures
No visual example yet
Explore the skillFrancis-Rings
We present StableAvatar, the first end-to-end video diffusion transformer, which synthesizes infinite-length high-quality audio-driven avatar videos without any post-pro…