Amphion
open-mmlab
Amphion (/æmˈfaɪən/) is a toolkit for Audio, Music, and Speech Generation. Its purpose is to support reproducible research and help junior researchers and engineers get…
OPENAGENTSKILL / DIRECTORY
Find a skill for your next task. Explore tools for Codex, Claude Code, Cursor and more.
Find a skill for your next task. Preview examples where available.
17–32 / 84
Results: 84
open-mmlab
Amphion (/æmˈfaɪən/) is a toolkit for Audio, Music, and Speech Generation. Its purpose is to support reproducible research and help junior researchers and engineers get…
snakers4
Silero Models: pre-trained text-to-speech models made embarrassingly simple
remsky
Dockerized FastAPI wrapper for Kokoro-82M text-to-speech model w/multiplatform CPU, AMD, NVIDIA GPU PyTorch support, handling, and auto-stitching
denizsafak
Generate audiobooks from EPUBs, PDFs and text with synchronized captions.
huggingface
A blazing fast inference solution for text embeddings models
TheJoeFin
Use OCR in Windows quickly and easily with Text Grab. With optional background process and notifications.
KoljaB
Converts text to speech in realtime
OpenMOSS
MOSS‑TTS Family is an open‑source speech and sound generation model family from MOSI.AI and the OpenMOSS team. It is designed for high‑fidelity, high‑expressiveness, and…
nateshmbhat
Offline Text To Speech synthesis for python
FunAudioLLM
Multilingual speech understanding: ASR + emotion recognition + audio event detection. 50+ languages, 15x faster than Whisper, non-autoregressive.
SublimeText-Markdown
Powerful Markdown package for Sublime Text with better syntax understanding and good color schemes.
espnet
End-to-End Speech Processing Toolkit
argmaxinc
On-device Speech AI for Apple Silicon
wenet-e2e
Production First and Production Ready End-to-End Speech Recognition Toolkit
cmusphinx
A small speech recognizer
Textualize
Rich is a Python library for rich text and beautiful formatting in the terminal.