Amphion
open-mmlab
Amphion (/æmˈfaɪən/) is a toolkit for Audio, Music, and Speech Generation. Its purpose is to support reproducible research and help junior researchers and engineers get…
OPENAGENTSKILL / DIRECTORY
Find a skill for your next task. Explore tools for Codex, Claude Code, Cursor and more.
Search retrieves candidates across the registry and ranks a bounded shortlist by task fit. This count is matching candidates, not the registry total. No suitable match? Try a specific tool or task.
65–80 / 114
Results: 114
open-mmlab
Amphion (/æmˈfaɪən/) is a toolkit for Audio, Music, and Speech Generation. Its purpose is to support reproducible research and help junior researchers and engineers get…
leejet
Diffusion model(SD,Flux,Wan,Qwen Image,Z-Image,...) inference in pure C/C++
snakers4
Silero Models: pre-trained text-to-speech models made embarrassingly simple
nl8590687
A Deep-Learning-Based Chinese Speech Recognition System 基于深度学习的中文语音识别系统
Pluviobyte
Build a reproducible local rough-cut workflow for talking-head or narrated screen-recording videos with transcript review gates.
modelscope
Open-source, accurate and easy-to-use video speech recognition & clipping tool. LLM-based AI clipping integrated.
wenet-e2e
Production First and Production Ready End-to-End Speech Recognition Toolkit
vllm-project
A framework for efficient model inference with omni-modality models
remsky
Dockerized FastAPI wrapper for Kokoro-82M text-to-speech model w/multiplatform CPU, AMD, NVIDIA GPU PyTorch support, handling, and auto-stitching
Breakthrough
:movie_camera: Python and OpenCV-based scene cut/transition detection program & library.
denizsafak
Generate audiobooks from EPUBs, PDFs and text with synchronized captions.
Picovoice
On-device wake word detection powered by deep learning
Owner-curated external sources. Not filtered by the scores or compatibility controls above; excluded from GitHub rankings and automatic installation.
No external entries match this query.