My Translator
phuc-nt
Real-time speech translation — macOS & Windows, free TTS, no server, your API keys only
OPENAGENTSKILL / DIRECTORY
Find a skill for your next task. Explore tools for Codex, Claude Code, Cursor and more.
Find a skill for your next task. Preview examples where available.
Page 7 · 16 shown · 743 public entries
Results: 743
phuc-nt
Real-time speech translation — macOS & Windows, free TTS, no server, your API keys only
janvarev
Ирина - русский голосовой ассистент для работы оффлайн. Поддерживает скиллы через плагины.
etro-js
Typescript video-editing library for the browser
lhotse-speech
Tools for handling multimodal data in machine learning projects.
diodiogod
A ComfyUI custom node integration for local multi-engine multi-language Text-to-Speech and Voice Conversion. Supports: RVC, Echo-TTS, Qwen3-TTS, Cozy Voice 3, Step Audio…
libAudioFlux
A library for audio and music analysis, feature extraction.
rsxdalv
A single Gradio + React WebUI with extensions for ACE-Step, OmniVoice, Kimi Audio, Piper TTS, GPT-SoVITS, CosyVoice, XTTSv2, DIA, Kokoro, OpenVoice, ParlerTTS, Stable Au…
JimmySadek
Claude Code skill: turn YouTube videos into structured, Obsidian-ready Markdown notes with full metadata, chapters, and transcripts
stemrollerapp
Isolate vocals, drums, bass, and other instrumental stems from any song
qxresearch
Python hands on tutorial with 50+ Python Application (10 lines of code) By @xiaowuc2
pndurette
Python library and CLI tool to interface with Google Translate's text-to-speech API
milvus-io
Dealing with all unstructured data, such as reverse image search, audio search, molecular search, video analysis, question and answer systems, NLP, etc.
iver56
A Python library for audio data augmentation. Useful for making audio ML models work well in the real world, not just in the lab.
hkchengrex
[CVPR 2025] MMAudio: Taming Multimodal Joint Training for High-Quality Video-to-Audio Synthesis
nexu-io
Audio generation skill — jingles, beds, voiceover, and sound effects. Routes music requests to Suno V5 / Udio / Lyria, speech to MiniMax TTS / FishAudio / ElevenLabs V3,…
FireRedTeam
Open-source industrial-grade ASR models supporting Mandarin, Chinese dialects and English, achieving a new SOTA on public Mandarin ASR benchmarks, while also offering ou…