No visual example yet
Explore the skillDsh Vision Toolkit
Anionex
让纯文本模型更好地做视觉任务的DeepSeek Harness插件:带意图的图片问答、长截图 OCR、UI 还原等|DeepSeek Harness-native integration for agent-vision-toolkit: image Q&A, long-screenshot OCR, UI restoration, g…
OPENAGENTSKILL / DIRECTORY
次のタスクに合うスキルを。Codex、Claude Code、Cursor などのツールを探せます。
171 Skills
検索結果: 171
No visual example yet
Explore the skillAnionex
让纯文本模型更好地做视觉任务的DeepSeek Harness插件:带意图的图片问答、长截图 OCR、UI 还原等|DeepSeek Harness-native integration for agent-vision-toolkit: image Q&A, long-screenshot OCR, UI restoration, g…
No visual example yet
Explore the skillxtreme1-io
Xtreme1 is an all-in-one data labeling and annotation platform for multimodal data training and supports 3D LiDAR point cloud, image, and LLM.
No visual example yet
Explore the skilllessthanoptimal
Fast computer vision library for SFM, calibration, fiducials, tracking, image processing, and more.

enricoros
AI suite powered by state-of-the-art models and providing advanced AI/AGI functions. Includes AI personas, AGI functions, world-class Beam multi-model chats, text-to-ima…
View previews · 2Alisa0808
Turn one topic into a narrated Vox-style paper-collage explainer or ad video, from script through captions.
作例を見る · 3No visual example yet
Explore the skillsmixs
AI film director skills for Claude agents: cinematic dramaturgy (Murch, blocking, montage) + exact prompt syntax for Seedance 2.5, Kling 3.0 Turbo/Omni, Veo 3.1, Nano Ba…
No visual example yet
Explore the skillgnipbao
Agent skill: convert Chinese story copy or ordered images into a hand-drawn diary-comic animation (silent MP4 picture track).
No visual example yet
Explore the skillPaddlePaddle
Turn any PDF or image document into structured data for your AI. A powerful, lightweight OCR toolkit that bridges the gap between images/PDFs and LLMs. Supports 100+ lan…
No visual example yet
Explore the skillopenai
CLIP (Contrastive Language-Image Pretraining), Predict the most relevant text snippet given an image
No visual example yet
Explore the skillyzhao062
A Python library for anomaly detection across tabular, time series, graph, text, and image data. 60+ detectors, benchmark-backed ADEngine orchestration, and an agentic w…
No visual example yet
Explore the skillNVlabs
SANA: Efficient High-Resolution Image Synthesis with Linear Diffusion Transformer
No visual example yet
Explore the skillbytedance
The official repo for “Dolphin: Document Image Parsing via Heterogeneous Anchor Prompting”, ACL, 2025.
No visual example yet
Explore the skillemilkowalski
Reverse-lookup glossary that turns a vague description of a web animation or motion effect into its exact term ("the bouncy thing when a popover opens" → Pop in; "the iO…
No visual example yet
Explore the skillmylxsw
An APP that integrates mainstream large language models and image generation models, built with Flutter, with fully open-source code.
No visual example yet
Explore the skillcrazy-max
Receive notifications when an image is updated on a Docker registry
No visual example yet
Explore the skillMaaXYZ
基于图像识别的自动化黑盒测试框架 | An automation black-box testing framework based on image recognition