OPENAGENTSKILL / DIRECTORY
Find a skill for your next task. Explore tools for Codex, Claude Code, Cursor and more.
289–304 / 476
Candidates in this shortlist, not the full registry. GitHub stars belong to repositories, not individual skills.
Results: 476
k2-fsa
Speech-to-text, text-to-speech, speaker diarization, speech enhancement, source separation, and VAD using next-gen Kaldi with onnxruntime without Internet connection. Su…
Picovoice
On-device Speech-to-Intent engine powered by deep learning
FireRedTeam
A SOTA Industrial-Grade All-in-One ASR system with ASR, VAD, LID, and Punc modules. FireRedASR2 supports Chinese (Mandarin, 20+ dialects/accents), English, code-switchin…
FireRedTeam
An Open-Sourced LLM-empowered Foundation TTS System
githubharald
Connectionist Temporal Classification (CTC) decoding algorithms: best path, beam search, lexicon search, prefix search, and token passing. Implemented in Python.
echogarden-project
Cross-platform speech toolset, used from the command-line or as a Node.js library. Includes a variety of engines for speech synthesis, speech recognition, forced alignme…
githubharald
Connectionist Temporal Classification (CTC) decoder with dictionary and language model.
RVC-Boss
1 min voice data can also be used to train a good TTS model! (few shot voice cloning)
m-bain
WhisperX: Automatic Speech Recognition with Word-level Timestamps (& Diarization)
babysor
🚀Clone a voice in 5 seconds to generate arbitrary speech in real-time
espeak-ng
eSpeak NG is an open source speech synthesizer that supports more than hundred languages and accents.
yl4579
StyleTTS 2: Towards Human-Level Text-to-Speech through Style Diffusion and Adversarial Training with Large Speech Language Models
RHVoice
a free and open source speech synthesizer for Russian and other languages
index-tts
An Industrial-Level Controllable and Efficient Zero-Shot Text-To-Speech System