Wan2.2
Wan-Video
Wan: Open and Advanced Large-Scale Video Generative Models
OPENAGENTSKILL / DIRECTORY
Find a skill for your next task. Explore tools for Codex, Claude Code, Cursor and more.
17–32 / 335
Candidates in this shortlist, not the full registry. GitHub stars belong to repositories, not individual skills.
Results: 335
Wan-Video
Wan: Open and Advanced Large-Scale Video Generative Models
HKUDS
"ViMax: Agentic Video Generation (Director, Screenwriter, Producer, and Video Generator All-in-One)"
nari-labs
A TTS model capable of generating ultra-realistic dialogue in one pass.
myshell-ai
Instant voice cloning by MIT and MyShell. Audio foundation model.
Azure-Samples
Microsoft Text-to-Speech API sample code in several languages, part of Cognitive Services.
Uberi
Speech recognition module for Python, supporting several engines and APIs, online and offline.
rany2
Use Microsoft Edge's online text-to-speech service from Python WITHOUT needing Microsoft Edge or Windows or an API key
thu-ml
Official code of Motus: A Unified Latent Action World Model
VectorSpaceLab
OmniGen: Unified Image Generation. https://arxiv.org/pdf/2409.11340
snap-research
Code for Motion Representations for Articulated Animation paper
sdkcarlos
A voice control - voice commands - speech recognition and speech synthesis javascript library. Create your own siri,google now or cortana with Google Chrome within your…
ali-vilab
Code for SCIS-2025 Paper "UniAnimate: Taming Unified Video Diffusion Models for Consistent Human Image Animation".
espeak-ng
eSpeak NG is an open source speech synthesizer that supports more than hundred languages and accents.
KnpLabs
PHP library allowing thumbnail, snapshot or PDF generation from a url or a html page. Wrapper for wkhtmltopdf/wkhtmltoimage