Lance
bytedance
A 3B-active-parameter native unified multimodal model for image and video understanding, generation, and editing.
OPENAGENTSKILL / DIRECTORY
Find a skill for your next task. Explore tools for Codex, Claude Code, Cursor and more.
17–32 / 58
Candidates in this shortlist, not the full registry. GitHub stars belong to repositories, not individual skills.
Results: 58
bytedance
A 3B-active-parameter native unified multimodal model for image and video understanding, generation, and editing.
FoundationVision
[CVPR 2025 Oral]Infinity ∞ : Scaling Bitwise AutoRegressive Modeling for High-Resolution Image Synthesis
OpenGVLab
InternGPT (iGPT) is an open source demo platform where you can easily showcase your AI models. Now it supports DragGAN, ChatGPT, ImageBind, multimodal chat like GPT-4, S…
FireRedTeam
FireRed-Image-Edit is a powerful image editing foundation model achieving open-source state-of-the-art performance with precise instruction following, high-fidelity gene…
FoundationVision
Autoregressive Model Beats Diffusion: 🦙 Llama for Scalable Image Generation
thu-ml
Official code of Motus: A Unified Latent Action World Model
ali-vilab
Timestep Embedding Tells: It's Time to Cache for Video Diffusion Model
snap-research
Code for Motion Representations for Articulated Animation paper
Aratako
A Flow Matching-based Text-to-Speech Model with Emoji-driven Style Control
thu-ml
[ICML 2026] Official codebase for "Causal Forcing: Autoregressive Diffusion Distillation Done Right for High-Quality Real-Time Interactive Video Generation" & Causal For…
NVlabs
rCM & Causal-rCM: Leading and Unified Algorithms/Infrastructures for Bidirectional/Autoregressive Video Diffusion Distillation at Scale
RVC-Boss
1 min voice data can also be used to train a good TTS model! (few shot voice cloning)
lone-cloud
A desktop app for running Large Language Models locally.
Lakonik
Official implementation of AsymFlow, pi-Flow, GMFlow