No visual example yet
Explore the skillSpeech Recognition
Uberi
Speech recognition module for Python, supporting several engines and APIs, online and offline.
OPENAGENTSKILL / DIRECTORY
Finde den passenden Skill für deine nächste Aufgabe mit Codex, Claude Code, Cursor und mehr.
29 Skills
Ergebnisse: 29
No visual example yet
Explore the skillUberi
Speech recognition module for Python, supporting several engines and APIs, online and offline.
No visual example yet
Explore the skillm-bain
WhisperX: Automatic Speech Recognition with Word-level Timestamps (& Diarization)
No visual example yet
Explore the skillTalAter
💬 Speech recognition for your site
No visual example yet
Explore the skillcmusphinx
A small speech recognizer
No visual example yet
Explore the skillMahmoudAshraf97
Automatic Speech Recognition with Speaker Diarization based on OpenAI Whisper
No visual example yet
Explore the skillflashlight
Facebook AI Research's Automatic Speech Recognition Toolkit
No visual example yet
Explore the skilljianchang512
Voice Recognition to Text Tool / 一个离线运行的本地音视频转字幕工具,输出json、srt字幕、纯文字格式
No visual example yet
Explore the skillhuggingface
Distilled variant of Whisper for speech recognition. 6x faster, 50% smaller, within 1% word error rate.
No visual example yet
Explore the skillalphacep
Offline speech recognition for Android with Vosk library.
No visual example yet
Explore the skillageitgey
The world's simplest facial recognition api for Python and the command line
No visual example yet
Explore the skillbabysor
🚀Clone a voice in 5 seconds to generate arbitrary speech in real-time
No visual example yet
Explore the skillrany2
Use Microsoft Edge's online text-to-speech service from Python WITHOUT needing Microsoft Edge or Windows or an API key
No visual example yet
Explore the skillespeak-ng
eSpeak NG is an open source speech synthesizer that supports more than hundred languages and accents.
No visual example yet
Explore the skillKoljaB
Converts text to speech in realtime
No visual example yet
Explore the skillOpenMOSS
MOSS‑TTS Family is an open‑source speech and sound generation model family from MOSI.AI and the OpenMOSS team. It is designed for high‑fidelity, high‑expressiveness, and…
No visual example yet
Explore the skilljaywalnut310
VITS: Conditional Variational Autoencoder with Adversarial Learning for End-to-End Text-to-Speech