No visual example yet
Explore the skillLance
bytedance
A 3B-active-parameter native unified multimodal model for image and video understanding, generation, and editing.
OPENAGENTSKILL / DIRECTORY
Trouvez un skill pour votre prochaine tâche avec Codex, Claude Code, Cursor et plus encore.
5 Skills
Résultats: 5
No visual example yet
Explore the skillbytedance
A 3B-active-parameter native unified multimodal model for image and video understanding, generation, and editing.
No visual example yet
Explore the skillnvidia-cosmos
Cosmos-Transfer2.5, built on top of Cosmos-Predict2.5, produces high-quality world simulations conditioned on multiple spatial control inputs.
No visual example yet
Explore the skillunum-cloud
Pocket-Sized Multimodal AI for content understanding and generation across multilingual texts, images, and 🔜 video, up to 5x faster than OpenAI CLIP and LLaVA 🖼️ & 🖋️
No visual example yet
Explore the skillmenyifang
Official implementation of "MIMO: Controllable Character Video Synthesis with Spatial Decomposed Modeling"
No visual example yet
Explore the skillTIGER-AI-Lab
Official Repo for "TheoremExplainAgent: Towards Video-based Multimodal Explanations for LLM Theorem Understanding" [ACL 2025 oral]