PaddleSpeech
PaddlePaddle
Easy-to-use Speech Toolkit including Self-Supervised Learning model, SOTA/Streaming ASR with punctuation, Streaming TTS with text frontend, Speaker Verification System,…
OPENAGENTSKILL / DIRECTORY
Find a skill for your next task. Explore tools for Codex, Claude Code, Cursor and more.
1–16 / 107
Candidates in this shortlist, not the full registry. GitHub stars belong to repositories, not individual skills.
Results: 107
PaddlePaddle
Easy-to-use Speech Toolkit including Self-Supervised Learning model, SOTA/Streaming ASR with punctuation, Streaming TTS with text frontend, Speaker Verification System,…
Uberi
Speech recognition module for Python, supporting several engines and APIs, online and offline.
nl8590687
A Deep-Learning-Based Chinese Speech Recognition System 基于深度学习的中文语音识别系统
hashicorp
Terraform provider for Azure Resource Manager
huggingface
Build local voice agents with open-source models
lenML
🍦 Speech-AI-Forge is a project developed around TTS generation model, implementing an API Server and a Gradio-based WebUI.
lhotse-speech
Tools for handling multimodal data in machine learning projects.
m-bain
WhisperX: Automatic Speech Recognition with Word-level Timestamps (& Diarization)
microsoft
A highly-customizable web-based client for Azure Bot Services.
FunAudioLLM
Multilingual speech understanding: ASR + emotion recognition + audio event detection. 50+ languages, 15x faster than Whisper, non-autoregressive.
alphacep
Offline speech recognition API for Android, iOS, Raspberry Pi and servers with Python, Java, C# and Node
argmaxinc
On-device Speech AI for Apple Silicon
Blaizzy
A text-to-speech (TTS), speech-to-text (STT) and speech-to-speech (STS) library built on Apple's MLX framework, providing efficient speech analysis on Apple Silicon.