Automatic Speech Recognition
zzw922cn
End-to-end Automatic Speech Recognition for Madarian and English in Tensorflow
OPENAGENTSKILL / DIRECTORY
Find a skill for your next task. Explore tools for Codex, Claude Code, Cursor and more.
1–16 / 16
Candidates in this shortlist, not the full registry. GitHub stars belong to repositories, not individual skills.
Results: 16
zzw922cn
End-to-end Automatic Speech Recognition for Madarian and English in Tensorflow
wenet-e2e
Production First and Production Ready End-to-End Speech Recognition Toolkit
FireRedTeam
Open-source industrial-grade ASR models supporting Mandarin, Chinese dialects and English, achieving a new SOTA on public Mandarin ASR benchmarks, while also offering ou…
R3gm
Synchronized Translation for Videos. Video dubbing
coqui-ai
🐸STT - The deep learning toolkit for Speech-to-Text. Training and deploying STT models has never been so easy.
varunshenoy
An extensible, easy-to-use, and portable diffusion web UI 👨🎨
AutoArk
[AutoArk] GPA (General Purpose Audio) can do ASR, TTS and voice conversion with one tiny model!
Picovoice
On-device streaming speech-to-text engine powered by deep learning
FireRedTeam
A SOTA Industrial-Grade All-in-One ASR system with ASR, VAD, LID, and Punc modules. FireRedASR2 supports Chinese (Mandarin, 20+ dialects/accents), English, code-switchin…
m-bain
WhisperX: Automatic Speech Recognition with Word-level Timestamps (& Diarization)
MahmoudAshraf97
Automatic Speech Recognition with Speaker Diarization based on OpenAI Whisper
flashlight
Facebook AI Research's Automatic Speech Recognition Toolkit
byjlw
Analyze videos using LLMs, Computer Vision and Automatic Speech Recognition