No visual example yet
Explore the skillSTT
coqui-ai
🐸STT - The deep learning toolkit for Speech-to-Text. Training and deploying STT models has never been so easy.
OPENAGENTSKILL / DIRECTORY
Find a skill for your next task. Explore tools for Codex, Claude Code, Cursor and more.
Find a skill for your next task. Preview examples where available.
33–48 / 103
Results: 103
No visual example yet
Explore the skillcoqui-ai
🐸STT - The deep learning toolkit for Speech-to-Text. Training and deploying STT models has never been so easy.
No visual example yet
Explore the skillk2-fsa
Real-time speech recognition and voice activity detection (VAD) using next-gen Kaldi with ncnn without Internet connection. Support iOS, Android, Linux, macOS, Windows,…
No visual example yet
Explore the skillyeyupiaoling
Fine-tune the Whisper speech recognition model to support training without timestamp data, training with timestamp data, and training without speech data. Accelerate inf…
No visual example yet
Explore the skillmravanelli
pytorch-kaldi is a project for developing state-of-the-art DNN/RNN hybrid speech recognition systems. The DNN part is managed by pytorch, while feature extraction, label…
No visual example yet
Explore the skillnobody132
中文语音识别; Mandarin Automatic Speech Recognition;
No visual example yet
Explore the skillsooftware
[Unofficial] PyTorch implementation of "Conformer: Convolution-augmented Transformer for Speech Recognition" (INTERSPEECH 2020)
No visual example yet
Explore the skillalphacep
Offline speech recognition for Android with Vosk library.
No visual example yet
Explore the skillbabysor
🚀Clone a voice in 5 seconds to generate arbitrary speech in real-time
No visual example yet
Explore the skillk2-fsa
Speech-to-text server framework with next-gen Kaldi
No visual example yet
Explore the skillBinWang28
The hub for audio AI research: papers, open models, benchmarks & datasets across audio LLMs, speech recognition, TTS, music & audio generation.
No visual example yet
Explore the skillsandrohanea
Whisper.net. Speech to text made simple using Whisper Models
No visual example yet
Explore the skillVRCWizard
Speech to Text to Speech. Song now playing. Sends text as OSC messages to VRChat to display on avatar. (STTTS) (Speech to TTS) (VRC STT System) (VTuber TTS)
No visual example yet
Explore the skillPicovoice
On-device Speech-to-Intent engine powered by deep learning
No visual example yet
Explore the skillPicovoice
On-device streaming speech-to-text engine powered by deep learning
No visual example yet
Explore the skillFireRedTeam
A SOTA Industrial-Grade All-in-One ASR system with ASR, VAD, LID, and Punc modules. FireRedASR2 supports Chinese (Mandarin, 20+ dialects/accents), English, code-switchin…
No visual example yet
Explore the skillOpenBMB
VoxCPM2: Tokenizer-Free TTS for Multilingual Speech Generation, Creative Voice Design, and True-to-Life Cloning