SpeechT5
microsoft
Unified-Modal Speech-Text Pre-Training for Spoken Language Processing
OPENAGENTSKILL / DIRECTORY
Find a skill for your next task. Explore tools for Codex, Claude Code, Cursor and more.
Find a skill for your next task. Preview examples where available.
1–16 / 113
Results: 113
microsoft
Unified-Modal Speech-Text Pre-Training for Spoken Language Processing
PaddlePaddle
Easy-to-use Speech Toolkit including Self-Supervised Learning model, SOTA/Streaming ASR with punctuation, Streaming TTS with text frontend, Speaker Verification System,…
speechbrain
Uberi
Speech recognition module for Python, supporting several engines and APIs, online and offline.
nl8590687
A Deep-Learning-Based Chinese Speech Recognition System 基于深度学习的中文语音识别系统
zzw922cn
End-to-end Automatic Speech Recognition for Madarian and English in Tensorflow
🎙Speech recognition using the tensorflow deep learning framework, sequence-to-sequence neural networks
MiteshPuthran
The neural network model is capable of detecting five different male/female emotions from audio speeches. (Deep Learning, NLP, Python)
soniqo
AI speech toolkit for Apple Silicon — ASR, TTS, speech-to-speech, VAD, and diarization powered by MLX and CoreML
jamsch
Speech Recognition for React Native Expo projects
huggingface
Build local voice agents with open-source models
lenML
🍦 Speech-AI-Forge is a project developed around TTS generation model, implementing an API Server and a Gradio-based WebUI.
Azure-Samples
Microsoft Text-to-Speech API sample code in several languages, part of Cognitive Services.
lhotse-speech
Tools for handling multimodal data in machine learning projects.
Aivis-Project
AivisSpeech: AI Voice Imitation System - Text to Speech Software
julius-speech
Open-Source Large Vocabulary Continuous Speech Recognition Engine