ViMax
HKUDS
"ViMax: Agentic Video Generation (Director, Screenwriter, Producer, and Video Generator All-in-One)"
OPENAGENTSKILL / DIRECTORY
Find a skill for your next task. Explore tools for Codex, Claude Code, Cursor and more.
17–32 / 215
Candidates in this shortlist, not the full registry. GitHub stars belong to repositories, not individual skills.
Results: 215
HKUDS
"ViMax: Agentic Video Generation (Director, Screenwriter, Producer, and Video Generator All-in-One)"
nari-labs
A TTS model capable of generating ultra-realistic dialogue in one pass.
myshell-ai
Instant voice cloning by MIT and MyShell. Audio foundation model.
Uberi
Speech recognition module for Python, supporting several engines and APIs, online and offline.
rany2
Use Microsoft Edge's online text-to-speech service from Python WITHOUT needing Microsoft Edge or Windows or an API key
zai-org
text and image to video generation: CogVideoX (2024) and CogVideo (ICLR 2023)
gitmylo
A webui for different audio related Neural Networks
nl8590687
A Deep-Learning-Based Chinese Speech Recognition System 基于深度学习的中文语音识别系统
Breakthrough
:movie_camera: Python and OpenCV-based scene cut/transition detection program & library.
hao-ai-lab
A unified inference and post-training framework for accelerated video generation.
OpenMOSS
MOSS‑TTS Family is an open‑source speech and sound generation model family from MOSI.AI and the OpenMOSS team. It is designed for high‑fidelity, high‑expressiveness, and…
Tencent-Hunyuan
HunyuanVideo-1.5: A leading lightweight video generation model
netease-youdao
EmotiVoice 😊: a Multi-Voice and Prompt-Controlled TTS Engine