PaddleSpeech
PaddlePaddle
Easy-to-use Speech Toolkit including Self-Supervised Learning model, SOTA/Streaming ASR with punctuation, Streaming TTS with text frontend, Speaker Verification System,…
OPENAGENTSKILL / DIRECTORY
Find a skill for your next task. Explore tools for Codex, Claude Code, Cursor and more.
1–16 / 109
Results: 109
PaddlePaddle
Easy-to-use Speech Toolkit including Self-Supervised Learning model, SOTA/Streaming ASR with punctuation, Streaming TTS with text frontend, Speaker Verification System,…
Uberi
Speech recognition module for Python, supporting several engines and APIs, online and offline.
nl8590687
A Deep-Learning-Based Chinese Speech Recognition System 基于深度学习的中文语音识别系统
zzw922cn
End-to-end Automatic Speech Recognition for Madarian and English in Tensorflow
soniqo
AI speech toolkit for Apple Silicon — ASR, TTS, speech-to-speech, VAD, and diarization powered by MLX and CoreML
huggingface
Build local voice agents with open-source models
phuc-nt
Real-time speech translation — macOS & Windows, free TTS, no server, your API keys only
lenML
🍦 Speech-AI-Forge is a project developed around TTS generation model, implementing an API Server and a Gradio-based WebUI.
Azure-Samples
Microsoft Text-to-Speech API sample code in several languages, part of Cognitive Services.
mkiol
Speech Note Linux app. Note taking, reading and translating with offline Speech to Text, Text to Speech and Machine translation.
lhotse-speech
Tools for handling multimodal data in machine learning projects.
sandrohanea
Whisper.net. Speech to text made simple using Whisper Models
abus-aikorea
Gradio WebUI for creators and developers, featuring key TTS (Edge-TTS, kokoro) and zero-shot Voice Cloning (E2 & F5-TTS, CosyVoice), with Whisper audio processing, YouTu…
m-bain
WhisperX: Automatic Speech Recognition with Word-level Timestamps (& Diarization)
Owner-curated external sources. Not filtered by the scores or compatibility controls above; excluded from GitHub rankings and automatic installation.
No external entries match this query.