Echomimic V2
antgroup
[CVPR 2025] EchoMimicV2: Towards Striking, Simplified, and Semi-Body Human Animation
OPENAGENTSKILL / DIRECTORY
Find a skill for your next task. Explore tools for Codex, Claude Code, Cursor and more.
1–9 / 9
Candidates in this shortlist, not the full registry. GitHub stars belong to repositories, not individual skills.
Results: 9
antgroup
[CVPR 2025] EchoMimicV2: Towards Striking, Simplified, and Semi-Body Human Animation
bytedance
[ICCV 2025] 🔥🔥 UNO: A Universal Customization Method for Both Single and Multi-Subject Conditioning
Tencent-Hunyuan
HunyuanCustom: A Multimodal-Driven Architecture for Customized Video Generation
invoke-ai
Invoke is a leading creative engine for Stable Diffusion models, empowering professionals, artists, and enthusiasts to generate and create visual media using the latest…
open-mmlab
Amphion (/æmˈfaɪən/) is a toolkit for Audio, Music, and Speech Generation. Its purpose is to support reproducible research and help junior researchers and engineers get…
mravanelli
pytorch-kaldi is a project for developing state-of-the-art DNN/RNN hybrid speech recognition systems. The DNN part is managed by pytorch, while feature extraction, label…
Francis-Rings
We present StableAvatar, the first end-to-end video diffusion transformer, which synthesizes infinite-length high-quality audio-driven avatar videos without any post-pro…
Aratako
A Flow Matching-based Text-to-Speech Model with Emoji-driven Style Control