Echomimic V2
antgroup
[CVPR 2025] EchoMimicV2: Towards Striking, Simplified, and Semi-Body Human Animation
OPENAGENTSKILL / DIRECTORY
Find a skill for your next task. Explore tools for Codex, Claude Code, Cursor and more.
113–128 / 173
Candidates in this shortlist, not the full registry. GitHub stars belong to repositories, not individual skills.
Results: 173
antgroup
[CVPR 2025] EchoMimicV2: Towards Striking, Simplified, and Semi-Body Human Animation
OpenMOSS
MOSS‑TTS Family is an open‑source speech and sound generation model family from MOSI.AI and the OpenMOSS team. It is designed for high‑fidelity, high‑expressiveness, and…
SkyworkAI
Matrix-Game 3.0: Real-Time and Streaming Interactive World Model with Long-Horizon Memory
fudan-generative-vision
[ECCV 2024] Champ: Controllable and Consistent Human Image Animation with 3D Parametric Guidance
open-mmlab
OpenMMLab Multimodal Advanced, Generative, and Intelligent Creation Toolbox. Unlock the magic 🪄: Generative-AI (AIGC), easy-to-use APIs, awsome model zoo, diffusion mod…
letoram
Arcan - [Display Server, Multimedia Framework, Game Engine] -> "Desktop Engine"
s60sc
ESP32 Camera motion capture application to record JPEGs to SD card as AVI files and stream to browser as MJPEG. If a microphone is installed then a WAV file is also crea…
OpenGVLab
InternGPT (iGPT) is an open source demo platform where you can easily showcase your AI models. Now it supports DragGAN, ChatGPT, ImageBind, multimodal chat like GPT-4, S…
kane50613
Render JSX, HTML, and CSS to images without a headless browser. OG cards, animated GIFs, and video frames from Node.js, edge runtimes, browsers, or Rust. Drop-in next/og…
nyanmisaka
FFmpeg with async and zero-copy Rockchip MPP & RGA support
thu-ml
[ICML2025] SpargeAttention: A training-free sparse attention that accelerates any model inference.
bytedance
🔥 [ICCV 2025 Highlight] InfiniteYou: Flexible Photo Recrafting While Preserving Your Identity
ali-vilab
Official implementations for paper: DreamTalk: When Expressive Talking Head Generation Meets Diffusion Probabilistic Models
ThioJoe
Automatically translates the text of a video based on a subtitle file, and then uses AI voice services to create a new dubbed & translated audio track where the speech i…
thu-ml
Official code of Motus: A Unified Latent Action World Model