Wenet
wenet-e2e
Production First and Production Ready End-to-End Speech Recognition Toolkit
OPENAGENTSKILL / DIRECTORY
Find a skill for your next task. Explore tools for Codex, Claude Code, Cursor and more.
273–288 / 478
Candidates in this shortlist, not the full registry. GitHub stars belong to repositories, not individual skills.
Results: 478
wenet-e2e
Production First and Production Ready End-to-End Speech Recognition Toolkit
vllm-project
A framework for efficient model inference with omni-modality models
denizsafak
Generate audiobooks from EPUBs, PDFs and text with synchronized captions.
Picovoice
On-device wake word detection powered by deep learning
phillipi
Image-to-image translation with conditional adversarial nets
mozilla
:robot: :speech_balloon: Deep learning for Text to Speech (Discussion forum: https://discourse.mozilla.org/c/tts)
hao-ai-lab
A unified inference and post-training framework for accelerated video generation.
IAHispano
A simple, high-quality voice conversion tool focused on ease of use and performance.
netease-youdao
EmotiVoice 😊: a Multi-Voice and Prompt-Controlled TTS Engine
PaddlePaddle
PaddlePaddle GAN library, including lots of interesting applications like First-Order motion transfer, Wav2Lip, picture repair, image editing, photo2cartoon, image style…
myshell-ai
High-quality multi-lingual text-to-speech library by MyShell.ai. Support English, Spanish, French, Chinese, Japanese and Korean.
thu-ml
TurboDiffusion: 100–200× Acceleration for Video Diffusion Models
ModelTC
Lightweight Image Video Action Generation Inference Framework
Owner-curated external sources. Not filtered by the scores or compatibility controls above; excluded from GitHub rankings and automatic installation.
No external entries match this query.