Pyramid Flow
jy0205
[ICLR 2025] Pyramidal Flow Matching for Efficient Video Generative Modeling
OPENAGENTSKILL / DIRECTORY
Find a skill for your next task. Explore tools for Codex, Claude Code, Cursor and more.
33β48 / 350
Candidates in this shortlist, not the full registry. GitHub stars belong to repositories, not individual skills.
Results: 350
jy0205
[ICLR 2025] Pyramidal Flow Matching for Efficient Video Generative Modeling
Doubiiu
[ECCV 2024, Oral] DynamiCrafter: Animating Open-domain Images with Video Diffusion Priors
TMElyralab
MuseV: Infinite-length and High Fidelity Virtual Human Video Generation with Visual Conditioned Parallel Denoising
boona13
Seamlessly extend any image in any direction with AI. Open-source web app powered by Gemini via OpenRouter, with Poisson-blended seams and best-of-3 variant picker.
PKU-YuanGroup
[TPAMI 2025π₯] MagicTime: Time-lapse Video Generation Models as Metamorphic Simulators
Phantom-video
Phantom: Subject-Consistent Video Generation via Cross-Modal Alignment
thu-ml
[ICML2025] SpargeAttention: A training-free sparse attention that accelerates any model inference.
aqm857886159
Open-source, local-first desktop app for AI video creation: write a script β generate images & video β edit on a timeline β export. Bring your own model & API key β everβ¦
Francis-Rings
We present StableAvatar, the first end-to-end video diffusion transformer, which synthesizes infinite-length high-quality audio-driven avatar videos without any post-proβ¦
Finrandojin
AI-powered multi-voice audiobook generator β LLM script annotation, voice cloning, voice design, LoRA training, per-line style control, and export to MP3, chaptered M4B,β¦
ali-vilab
Official implementations for paper: DreamTalk: When Expressive Talking Head Generation Meets Diffusion Probabilistic Models
Tencent-Hunyuan
HunyuanCustom: A Multimodal-Driven Architecture for Customized Video Generation
wladradchenko
Wunjo CE: Face Swap, Lip Sync, Control Remove Objects & Text & Background, Restyling, Audio Separator, Clone Voice, Video Generation. Open Source, Local & Free.
thu-ml
Official code of Motus: A Unified Latent Action World Model
menyifang
Official implementation of "MIMO: Controllable Character Video Synthesis with Spatial Decomposed Modeling"
lucidrains
Implementation of Video Diffusion Models, Jonathan Ho's new paper extending DDPMs to Video Generation - in Pytorch