Kaldi
kaldi-asr
kaldi-asr/kaldi is the official location of the Kaldi project.
OPENAGENTSKILL / DIRECTORY
Find a skill for your next task. Explore tools for Codex, Claude Code, Cursor and more.
321–336 / 476
Candidates in this shortlist, not the full registry. GitHub stars belong to repositories, not individual skills.
Results: 476
kaldi-asr
kaldi-asr/kaldi is the official location of the Kaldi project.
argmaxinc
On-device Speech AI for Apple Silicon
microsoft
Unified-Modal Speech-Text Pre-Training for Spoken Language Processing
Blaizzy
A text-to-speech (TTS), speech-to-text (STT) and speech-to-speech (STS) library built on Apple's MLX framework, providing efficient speech analysis on Apple Silicon.
jishengpeng
[ICLR 2025] SOTA discrete acoustic codec models with 40/75 tokens per second for audio language modeling
open-mmlab
Amphion (/æmˈfaɪən/) is a toolkit for Audio, Music, and Speech Generation. Its purpose is to support reproducible research and help junior researchers and engineers get…
snakers4
Silero Models: pre-trained text-to-speech models made embarrassingly simple
nl8590687
A Deep-Learning-Based Chinese Speech Recognition System 基于深度学习的中文语音识别系统
TensorSpeech
:zap: TensorFlowASR: Almost State-of-the-art Automatic Speech Recognition in Tensorflow 2. Supported languages that can use characters or subwords
wenet-e2e
Production First and Production Ready End-to-End Speech Recognition Toolkit
remsky
Dockerized FastAPI wrapper for Kokoro-82M text-to-speech model w/multiplatform CPU, AMD, NVIDIA GPU PyTorch support, handling, and auto-stitching
denizsafak
Generate audiobooks from EPUBs, PDFs and text with synchronized captions.
Picovoice
On-device wake word detection powered by deep learning