HunyuanVideo
Tencent-Hunyuan
HunyuanVideo: A Systematic Framework For Large Video Generation Model
OPENAGENTSKILL / DIRECTORY
Find a skill for your next task. Explore tools for Codex, Claude Code, Cursor and more.
33–48 / 351
Candidates in this shortlist, not the full registry. GitHub stars belong to repositories, not individual skills.
Results: 351
Tencent-Hunyuan
HunyuanVideo: A Systematic Framework For Large Video Generation Model
MoonInTheRiver
DiffSinger: Singing Voice Synthesis via Shallow Diffusion Mechanism (SVS & TTS); AAAI 2022; Official code
sanchit-gandhi
JAX implementation of OpenAI's Whisper model for up to 70x speed-up on TPU.
pydn
A powerful tool that translates ComfyUI workflows into executable Python code.
jianchang512
Translate the video from one language to another and embed dubbing & subtitles.
Doubiiu
[ECCV 2024, Oral] DynamiCrafter: Animating Open-domain Images with Video Diffusion Priors
cubist38
A high-performance API server that provides OpenAI-compatible endpoints for MLX models. Developed using Python and powered by the FastAPI framework, it provides an effic…
marytts
MARY TTS -- an open-source, multilingual text-to-speech synthesis system written in pure java
Xeron2000
故事想法 → 多智能体协作 → 漫剧成片 | 基于 LangGraph 的 AI 漫剧生成平台
FireRedTeam
FireRed-Image-Edit is a powerful image editing foundation model achieving open-source state-of-the-art performance with precise instruction following, high-fidelity gene…
HKUDS
"ViMax: Agentic Video Generation (Director, Screenwriter, Producer, and Video Generator All-in-One)"
Azure-Samples
Microsoft Text-to-Speech API sample code in several languages, part of Cognitive Services.
julius-speech
Open-Source Large Vocabulary Continuous Speech Recognition Engine
Uberi
Speech recognition module for Python, supporting several engines and APIs, online and offline.