No visual example yet
Explore the skillViMax
HKUDS
"ViMax: Agentic Video Generation (Director, Screenwriter, Producer, and Video Generator All-in-One)"
OPENAGENTSKILL / DIRECTORY
Find a skill for your next task. Explore tools for Codex, Claude Code, Cursor and more.
115 Skills
Results: 115
No visual example yet
Explore the skillHKUDS
"ViMax: Agentic Video Generation (Director, Screenwriter, Producer, and Video Generator All-in-One)"
No visual example yet
Explore the skillduixcom
🚀 Truly open-source AI avatar(digital human) toolkit for offline video generation and digital human cloning.
No visual example yet
Explore the skillespnet
End-to-End Speech Processing Toolkit
No visual example yet
Explore the skillgyroflow
Video stabilization using gyroscope data
No visual example yet
Explore the skillUberi
Speech recognition module for Python, supporting several engines and APIs, online and offline.
No visual example yet
Explore the skillrany2
Use Microsoft Edge's online text-to-speech service from Python WITHOUT needing Microsoft Edge or Windows or an API key
No visual example yet
Explore the skillNVlabs
SANA: Efficient High-Resolution Image Synthesis with Linear Diffusion Transformer
No visual example yet
Explore the skilldebpalash
The open-source ElevenLabs alternative for local voice cloning, design, create, dubbing and dictation Desktop App
No visual example yet
Explore the skillBlaizzy
A text-to-speech (TTS), speech-to-text (STT) and speech-to-speech (STS) library built on Apple's MLX framework, providing efficient speech analysis on Apple Silicon.
No visual example yet
Explore the skillTalAter
💬 Speech recognition for your site
No visual example yet
Explore the skillargmaxinc
On-device Speech AI for Apple Silicon
No visual example yet
Explore the skillleejet
Diffusion model(SD,Flux,Wan,Qwen Image,Z-Image,...) inference in pure C/C++
No visual example yet
Explore the skillsnakers4
Silero Models: pre-trained text-to-speech models made embarrassingly simple
No visual example yet
Explore the skillmodelscope
Open-source, accurate and easy-to-use video speech recognition & clipping tool. LLM-based AI clipping integrated.
No visual example yet
Explore the skillwenet-e2e
Production First and Production Ready End-to-End Speech Recognition Toolkit
No visual example yet
Explore the skillvllm-project
A framework for efficient model inference with omni-modality models