Silero Models
snakers4
Silero Models: pre-trained text-to-speech models made embarrassingly simple
OPENAGENTSKILL / DIRECTORY
Find a skill for your next task. Explore tools for Codex, Claude Code, Cursor and more.
1–16 / 68
Candidates in this shortlist, not the full registry. GitHub stars belong to repositories, not individual skills.
Results: 68
snakers4
Silero Models: pre-trained text-to-speech models made embarrassingly simple
huggingface
🤗 Diffusers: State-of-the-art diffusion models for image, video, and audio generation in PyTorch.
Tencent-Hunyuan
HunyuanVideo: A Systematic Framework For Large Video Generation Model
openvinotoolkit
OpenVINO™ is an open source toolkit for optimizing and deploying AI inference
NVlabs
SANA: Efficient High-Resolution Image Synthesis with Linear Diffusion Transformer
NVIDIA
State-of-the-Art Deep Learning scripts organized by models - easy to train and deploy with reproducible accuracy and performance on enterprise-grade infrastructure.
wenet-e2e
Production First and Production Ready End-to-End Speech Recognition Toolkit
FoundationVision
[NeurIPS 2024 Best Paper Award][GPT beats diffusion🔥] [scaling laws in visual generation📈] Official impl. of "Visual Autoregressive Modeling: Scalable Image Generation…
hao-ai-lab
A unified inference and post-training framework for accelerated video generation.
open-mmlab
OpenMMLab Multimodal Advanced, Generative, and Intelligent Creation Toolbox. Unlock the magic 🪄: Generative-AI (AIGC), easy-to-use APIs, awsome model zoo, diffusion mod…
thu-ml
TurboDiffusion: 100–200× Acceleration for Video Diffusion Models
ModelTC
Lightweight Image Video Action Generation Inference Framework
yl4579
StyleTTS 2: Towards Human-Level Text-to-Speech through Style Diffusion and Adversarial Training with Large Speech Language Models
PKU-YuanGroup
Helios: Real Real-Time Long Video Generation Model
bytedance
SALMONN family: A suite of advanced multi-modal LLMs