No visual example yet
Explore the skillazure-ai-vision
MicrosoftDocs
Expert knowledge for Azure AI Vision development including decision making, limits & quotas, configuration, integrations & coding patterns, and deployment. Use when usin…
OPENAGENTSKILL / DIRECTORY
Find a skill for your next task. Explore tools for Codex, Claude Code, Cursor and more.
21 Skills
Results: 21
No visual example yet
Explore the skillMicrosoftDocs
Expert knowledge for Azure AI Vision development including decision making, limits & quotas, configuration, integrations & coding patterns, and deployment. Use when usin…
No visual example yet
Explore the skillhyperb1iss
Use this skill when an agent must generate or edit raster images through Codex, especially from Claude Code or another harness without native image generation. Activates…
No visual example yet
Explore the skillbytedance
A 3B-active-parameter native unified multimodal model for image and video understanding, generation, and editing.
No visual example yet
Explore the skillJIA-Lab-research
This project is the official implementation of 'DreamOmni2: Multimodal Instruction-based Editing and Generation (CVPR2026 Highlight)''
No visual example yet
Explore the skillFireRedTeam
FireRed-Image-Edit is a powerful image editing foundation model achieving open-source state-of-the-art performance with precise instruction following, high-fidelity gene…
No visual example yet
Explore the skillOpenGVLab
Unofficial Implementation of DragGAN - "Drag Your GAN: Interactive Point-based Manipulation on the Generative Image Manifold" (DragGAN 全功能实现,在线Demo,本地部署试用,代码、模型已全部开源,支持W…
No visual example yet
Explore the skillali-vilab
Official implementations for paper: Anydoor: zero-shot object-level image customization
No visual example yet
Explore the skillbytedance
Bernini is a unified framework for video generation and editing that combines an MLLM-based semantic planner with a DiT-based renderer.
No visual example yet
Explore the skillermongroup
PyTorch implementation for SDEdit: Image Synthesis and Editing with Stochastic Differential Equations
No visual example yet
Explore the skillShilin-LU
[ICLR 2025] "Robust Watermarking Using Generative Priors Against Image Editing: From Benchmarking to Advances" (Official Implementation)
No visual example yet
Explore the skillnxnai
[SIGGRAPH Asia 25] Voost: A Unified and Scalable Diffusion Transformer for Bidirectional Virtual Try-On and Try-Off
No visual example yet
Explore the skillSOTAMak1r
A Unified Visual Generator with Interleaved OmniModal Context
No visual example yet
Explore the skillnv-tlabs
[ICLR 2026] ChronoEdit: Towards Temporal Reasoning for Image Editing and World Simulation

advimman
Official repository for the paper "High-Resolution Daytime Translation Without Domain Labels" (CVPR2020, Oral)
View previews · 1No visual example yet
Explore the skillinclusionAI
Code release for Ming-UniVision: Joint Image Understanding and Geneation with a Continuous Unified Tokenizer
No visual example yet
Explore the skillKevin-thu
Official Code for DiffMorpher: Unleashing the Capability of Diffusion Models for Image Morphing (CVPR 2024)
Design direction, Figma implementation, UI review, React performance, browser QA, and safe preview deployments.
Code review, repo inspection, testing, planning, shipping, and engineering workflows for coding agents.
Research, recent-events briefings, PDF parsing, markdown conversion, RAG ingestion, and knowledge workflows.
Stock analysis, market research, quant backtesting, financial data, and investment research skills.
Crawling, scraping, extraction, browser automation, structured data capture, and website-to-markdown workflows.
Presentation generation, editable PPTX, slide decks, speaker notes, and visual storytelling workflows.
Choose React motion graphics, generated footage or editing; compare dependencies and inspect actual outputs.
Image, video, creative production, UI design, multimodal generation, and visual workflow skills.