暂未收录效果图
查看技能说明ESearch
xushengfeng
截屏 离线OCR 搜索翻译 以图搜图 贴图 录屏 万向滚动截屏 屏幕翻译 Screenshot Offline OCR Search Translate Search for picture Paste the picture on the screen Screen recorder Omnidirectional scrolling…
OPENAGENTSKILL / DIRECTORY
为下一项任务找到合适的技能。探索适用于 Codex、Claude Code、Cursor 等 Agent 的工具。
67 Skills
搜索结果: 67
暂未收录效果图
查看技能说明xushengfeng
截屏 离线OCR 搜索翻译 以图搜图 贴图 录屏 万向滚动截屏 屏幕翻译 Screenshot Offline OCR Search Translate Search for picture Paste the picture on the screen Screen recorder Omnidirectional scrolling…
暂未收录效果图
查看技能说明withoutbg
暂未收录效果图
查看技能说明CookSleep
暂未收录效果图
查看技能说明bytedance
A 3B-active-parameter native unified multimodal model for image and video understanding, generation, and editing.
暂未收录效果图
查看技能说明JIA-Lab-research
This project is the official implementation of 'DreamOmni2: Multimodal Instruction-based Editing and Generation (CVPR2026 Highlight)''
暂未收录效果图
查看技能说明FireRedTeam
FireRed-Image-Edit is a powerful image editing foundation model achieving open-source state-of-the-art performance with precise instruction following, high-fidelity gene…
暂未收录效果图
查看技能说明calesthio
xAI Grok image and video generation guide covering authentication, endpoints, prompt structure, image editing, reference-image video, and async polling.
暂未收录效果图
查看技能说明PaddlePaddle
PaddlePaddle GAN library, including lots of interesting applications like First-Order motion transfer, Wav2Lip, picture repair, image editing, photo2cartoon, image style…
暂未收录效果图
查看技能说明OpenGVLab
Unofficial Implementation of DragGAN - "Drag Your GAN: Interactive Point-based Manipulation on the Generative Image Manifold" (DragGAN 全功能实现,在线Demo,本地部署试用,代码、模型已全部开源,支持W…
暂未收录效果图
查看技能说明ali-vilab
Official implementations for paper: Anydoor: zero-shot object-level image customization
暂未收录效果图
查看技能说明aden-hive
Required before calling image_generate. Create and edit images from a prompt — generate an image, make a picture / logo / illustration / icon / banner / poster / thumbna…
暂未收录效果图
查看技能说明OpenGVLab
InternGPT (iGPT) is an open source demo platform where you can easily showcase your AI models. Now it supports DragGAN, ChatGPT, ImageBind, multimodal chat like GPT-4, S…
暂未收录效果图
查看技能说明bytedance
Bernini is a unified framework for video generation and editing that combines an MLLM-based semantic planner with a DiT-based renderer.
暂未收录效果图
查看技能说明HorizonWind2004
[ICLR 2026] Official repo of paper "Reconstruction Alignment Improves Unified Multimodal Models". Unlocking the Massive Zero-shot Potential in Unified Multimodal Models…
暂未收录效果图
查看技能说明ermongroup
PyTorch implementation for SDEdit: Image Synthesis and Editing with Stochastic Differential Equations
暂未收录效果图
查看技能说明dreiachse-cyber
Local cockpit for Codex imagegen, pixel art, image editing, animation, and sprite-sheet workflows.