UniAnimate
ali-vilab
Code for SCIS-2025 Paper "UniAnimate: Taming Unified Video Diffusion Models for Consistent Human Image Animation".
OPENAGENTSKILL / DIRECTORY
Find a skill for your next task. Explore tools for Codex, Claude Code, Cursor and more.
257–272 / 478
Candidates in this shortlist, not the full registry. GitHub stars belong to repositories, not individual skills.
Results: 478
ali-vilab
Code for SCIS-2025 Paper "UniAnimate: Taming Unified Video Diffusion Models for Consistent Human Image Animation".
cure-lab
[ICLR24] Official implementation of the paper “MagicDrive: Street View Generation with Diverse 3D Geometry Control”
ermongroup
PyTorch implementation for SDEdit: Image Synthesis and Editing with Stochastic Differential Equations
shadow2496
Official PyTorch implementation of "VITON-HD: High-Resolution Virtual Try-On via Misalignment-Aware Normalization" (CVPR 2021)
zai-org
CogView4, CogView3-Plus and CogView3(ECCV 2024)
alumae
Real-time full-duplex speech recognition server, based on the Kaldi toolkit and the GStreamer framwork.
omerbt
Official Pytorch Implementation for "MultiDiffusion: Fusing Diffusion Paths for Controlled Image Generation" presenting "MultiDiffusion" (ICML 2023)
showlab
[ECCV 2024 Oral] MotionDirector: Motion Customization of Text-to-Video Diffusion Models.
videoflow
Python framework that facilitates the quick development of complex video analysis applications and other series-processing based applications in a multiprocessing enviro…
TensorSpeech
:zap: TensorFlowASR: Almost State-of-the-art Automatic Speech Recognition in Tensorflow 2. Supported languages that can use characters or subwords
Aratako
A Flow Matching-based Text-to-Speech Model with Emoji-driven Style Control
alesaccoia
Near-Realtime audio transcription using self-hosted Whisper and WebSocket in Python/JS
ManimCommunity
Manim plugin for all things voiceover
k2-fsa
Speech-to-text server framework with next-gen Kaldi