FunClip
modelscope
Open-source, accurate and easy-to-use video speech recognition & clipping tool. LLM-based AI clipping integrated.
OPENAGENTSKILL / DIRECTORY
Find a skill for your next task. Explore tools for Codex, Claude Code, Cursor and more.
Find a skill for your next task. Preview examples where available.
1–16 / 57
Results: 57
modelscope
Open-source, accurate and easy-to-use video speech recognition & clipping tool. LLM-based AI clipping integrated.
denizsafak
Generate audiobooks from EPUBs, PDFs and text with synchronized captions.
google-ai-edge
Cross-platform, customizable ML solutions for live and streaming media.
RVC-Boss
1 min voice data can also be used to train a good TTS model! (few shot voice cloning)
m-bain
WhisperX: Automatic Speech Recognition with Word-level Timestamps (& Diarization)
babysor
🚀Clone a voice in 5 seconds to generate arbitrary speech in real-time
ggml-org
Port of OpenAI's Whisper model in C/C++
jianchang512
Translate the video from one language to another and embed dubbing & subtitles.
openvinotoolkit
OpenVINO™ is an open source toolkit for optimizing and deploying AI inference
nari-labs
A TTS model capable of generating ultra-realistic dialogue in one pass.
Uberi
Speech recognition module for Python, supporting several engines and APIs, online and offline.
rany2
Use Microsoft Edge's online text-to-speech service from Python WITHOUT needing Microsoft Edge or Windows or an API key
TalAter
💬 Speech recognition for your site
wenet-e2e
Production First and Production Ready End-to-End Speech Recognition Toolkit
espeak-ng
eSpeak NG is an open source speech synthesizer that supports more than hundred languages and accents.
cmusphinx
A small speech recognizer