No visual example yet
Explore the skillMultimodal Garment Designer
aimagelab
This is the official repository for the paper "Multimodal Garment Designer: Human-Centric Latent Diffusion Models for Fashion Image Editing". ICCV 2023
OPENAGENTSKILL / DIRECTORY
Find a skill for your next task. Explore tools for Codex, Claude Code, Cursor and more.
Find a skill for your next task. Preview examples where available.
1–9 / 9
Results: 9
No visual example yet
Explore the skillaimagelab
This is the official repository for the paper "Multimodal Garment Designer: Human-Centric Latent Diffusion Models for Fashion Image Editing". ICCV 2023
No visual example yet
Explore the skillmlfoundations
An open-source framework for training large multimodal models.
No visual example yet
Explore the skillinvictus717
Meta-Transformer for Unified Multimodal Learning
No visual example yet
Explore the skillAutoArk
EVA OS — A real-time multimodal AIOS for next-generation hardware, enabling your devices being “alive” and as intelligent as a real brain.
No visual example yet
Explore the skillMMMU-Benchmark
This repo contains evaluation code for the paper "MMMU: A Massive Multi-discipline Multimodal Understanding and Reasoning Benchmark for Expert AGI"
No visual example yet
Explore the skillpliang279
[NeurIPS 2021] Multiscale Benchmarks for Multimodal Representation Learning
No visual example yet
Explore the skillzeyofu
This repo contains evaluation code for the paper "BLINK: Multimodal Large Language Models Can See but Not Perceive". https://arxiv.org/abs/2404.12390 [ECCV 2024]
No visual example yet
Explore the skillxtreme1-io
Xtreme1 is an all-in-one data labeling and annotation platform for multimodal data training and supports 3D LiDAR point cloud, image, and LLM.
No visual example yet
Explore the skillfoxglove
Multimodal visualization and data platform