PaddleSpeech
PaddlePaddle
Easy-to-use Speech Toolkit including Self-Supervised Learning model, SOTA/Streaming ASR with punctuation, Streaming TTS with text frontend, Speaker Verification System,…
OPENAGENTSKILL / DIRECTORY
Find a skill for your next task. Explore tools for Codex, Claude Code, Cursor and more.
Find a skill for your next task. Preview examples where available.
1–16 / 105
Results: 105
PaddlePaddle
Easy-to-use Speech Toolkit including Self-Supervised Learning model, SOTA/Streaming ASR with punctuation, Streaming TTS with text frontend, Speaker Verification System,…
speechbrain
Uberi
Speech recognition module for Python, supporting several engines and APIs, online and offline.
nl8590687
A Deep-Learning-Based Chinese Speech Recognition System 基于深度学习的中文语音识别系统
coqui-ai
🐸STT - The deep learning toolkit for Speech-to-Text. Training and deploying STT models has never been so easy.
zzw922cn
End-to-end Automatic Speech Recognition for Madarian and English in Tensorflow
🎙Speech recognition using the tensorflow deep learning framework, sequence-to-sequence neural networks
microsoft
Unified-Modal Speech-Text Pre-Training for Spoken Language Processing
MiteshPuthran
The neural network model is capable of detecting five different male/female emotions from audio speeches. (Deep Learning, NLP, Python)
coqui-ai
💎 A list of accessible speech corpora for ASR, TTS, and other Speech Technologies
ictnlp
StreamSpeech is an “All in One” seamless model for offline and simultaneous speech recognition, speech translation and speech synthesis.
soniqo
AI speech toolkit for Apple Silicon — ASR, TTS, speech-to-speech, VAD, and diarization powered by MLX and CoreML
jamsch
Speech Recognition for React Native Expo projects
huggingface
Build local voice agents with open-source models
lenML
🍦 Speech-AI-Forge is a project developed around TTS generation model, implementing an API Server and a Gradio-based WebUI.
Azure-Samples
Microsoft Text-to-Speech API sample code in several languages, part of Cognitive Services.