FastVideo
hao-ai-lab
A unified inference and post-training framework for accelerated video generation.
OPENAGENTSKILL / DIRECTORY
Find a skill for your next task. Explore tools for Codex, Claude Code, Cursor and more.
1–14 / 14
Candidates in this shortlist, not the full registry. GitHub stars belong to repositories, not individual skills.
Results: 14
hao-ai-lab
A unified inference and post-training framework for accelerated video generation.
yl4579
StyleTTS 2: Towards Human-Level Text-to-Speech through Style Diffusion and Adversarial Training with Large Speech Language Models
Finrandojin
AI-powered multi-voice audiobook generator — LLM script annotation, voice cloning, voice design, LoRA training, per-line style control, and export to MP3, chaptered M4B,…
microsoft
Unified-Modal Speech-Text Pre-Training for Spoken Language Processing
unslothai
Unsloth Studio is a web UI for training and running open models like Gemma 4, Qwen3.6, DeepSeek, gpt-oss locally.
supertone-inc
Lightning-Fast, On-Device, Multilingual TTS — running natively via ONNX.
shivammehta25
[ICASSP 2024] 🍵 Matcha-TTS: A fast TTS architecture with conditional flow matching
tnfe
A fast video processing library based on node.js (一个基于node.js的高速视频制作库)
DigitalPhonetics
Controllable and fast Text-to-Speech for over 7000 languages!
coqui-ai
🐸STT - The deep learning toolkit for Speech-to-Text. Training and deploying STT models has never been so easy.
yeyupiaoling
Fine-tune the Whisper speech recognition model to support training without timestamp data, training with timestamp data, and training without speech data. Accelerate inf…
thu-ml
[ICML2025] SpargeAttention: A training-free sparse attention that accelerates any model inference.
mit-han-lab
[NeurIPS 2020] Differentiable Augmentation for Data-Efficient GAN Training
EvolvingLMMs-Lab
A simple, unified multimodal models training engine. Lean, flexible, and built for hacking at scale.