No visual example yet
Explore the skillDsh Vision Toolkit
Anionex
让纯文本模型更好地做视觉任务的DeepSeek Harness插件:带意图的图片问答、长截图 OCR、UI 还原等|DeepSeek Harness-native integration for agent-vision-toolkit: image Q&A, long-screenshot OCR, UI restoration, g…
OPENAGENTSKILL / DIRECTORY
Find a skill for your next task. Explore tools for Codex, Claude Code, Cursor and more.
10 Skills
Results: 10
No visual example yet
Explore the skillAnionex
让纯文本模型更好地做视觉任务的DeepSeek Harness插件:带意图的图片问答、长截图 OCR、UI 还原等|DeepSeek Harness-native integration for agent-vision-toolkit: image Q&A, long-screenshot OCR, UI restoration, g…
No visual example yet
Explore the skillriddleling
An iOS OCR Server Using Apple’s Vision Framework
No visual example yet
Explore the skillprivatenumber
macOS CLI for OCR and searchable PDFs using Apple's Vision framework
No visual example yet
Explore the skillgetomni-ai
OCR & Document Extraction using vision models
No visual example yet
Explore the skillicereed
Use LLMs and LLM Vision (OCR) to handle paperless-ngx - Document Digitalization powered by AI
No visual example yet
Explore the skillemcf
Get clean data from tricky documents, powered by vision-language models ⚡
No visual example yet
Explore the skillczlonkowski
Handle files and binary data in n8n correctly. Use when working with files, images, PDFs, attachments, uploads or downloads, base64, vision/multimodal input, or when an…
No visual example yet
Explore the skillliustack
Plug-in vision for text-only models. Hard rule: when a file path or URL with an image extension (.png, .jpg, .jpeg, .webp, .gif, .heic, .heif) appears anywhere in the co…
No visual example yet
Explore the skilljamjamjon
A Rust library integrated with ONNXRuntime, providing a collection of Computer Vison and Vision-Language models such as YOLO, FastVLM, and more.
No visual example yet
Explore the skillOpenRaiser
📄 [Skill] Vision-in-the-loop LaTeX typesetting agent — auto-compile, render, diagnose, and fix paper layouts