No visual example yet
Explore the skillDsh Vision Toolkit
Anionex
让纯文本模型更好地做视觉任务的DeepSeek Harness插件:带意图的图片问答、长截图 OCR、UI 还原等|DeepSeek Harness-native integration for agent-vision-toolkit: image Q&A, long-screenshot OCR, UI restoration, g…
OPENAGENTSKILL / DIRECTORY
Find a skill for your next task. Explore tools for Codex, Claude Code, Cursor and more.
15 Skills
Results: 15
No visual example yet
Explore the skillAnionex
让纯文本模型更好地做视觉任务的DeepSeek Harness插件:带意图的图片问答、长截图 OCR、UI 还原等|DeepSeek Harness-native integration for agent-vision-toolkit: image Q&A, long-screenshot OCR, UI restoration, g…
No visual example yet
Explore the skillAnionex
给纯文本 LLM agent 装上眼睛:图片问答、OCR、截图分析、视觉定位等一套视觉工具箱 + skill,并可无缝接入 Codex、Claude Code、OpenCode、Pi | Give text-only LLM agents vision: image Q&A, OCR, screenshot understanding,…
No visual example yet
Explore the skillriddleling
An iOS OCR Server Using Apple’s Vision Framework
No visual example yet
Explore the skillprivatenumber
macOS CLI for OCR and searchable PDFs using Apple's Vision framework
No visual example yet
Explore the skillAgents365-ai
Generate draw.io diagrams from natural language — 6 presets, vision self-check + up to 5-round refinement, codebase-to-diagram, 10,000+ official shapes & 321 AI/LLM bran…
No visual example yet
Explore the skillgetomni-ai
OCR & Document Extraction using vision models
No visual example yet
Explore the skillA9T9
Ui.Vision Open-Source RPA Software with Computer Vision, OCR, Anthropic Computer Use/LLM. Selenium IDE import/export.
No visual example yet
Explore the skillicereed
Use LLMs and LLM Vision (OCR) to handle paperless-ngx - Document Digitalization powered by AI
No visual example yet
Explore the skillARahim3
Fine-tune LLMs on your Mac with Apple Silicon. SFT, DPO, GRPO, Vision, TTS, STT, Embedding, and OCR fine-tuning — natively on MLX. Unsloth-compatible API.
No visual example yet
Explore the skillemcf
Get clean data from tricky documents, powered by vision-language models ⚡
No visual example yet
Explore the skillreworkd
Vision utilities for web interaction agents 👀
No visual example yet
Explore the skillczlonkowski
Handle files and binary data in n8n correctly. Use when working with files, images, PDFs, attachments, uploads or downloads, base64, vision/multimodal input, or when an…
No visual example yet
Explore the skillliustack
Plug-in vision for text-only models. Hard rule: when a file path or URL with an image extension (.png, .jpg, .jpeg, .webp, .gif, .heic, .heif) appears anywhere in the co…
No visual example yet
Explore the skilljamjamjon
A Rust library integrated with ONNXRuntime, providing a collection of Computer Vison and Vision-Language models such as YOLO, FastVLM, and more.
No visual example yet
Explore the skillOpenRaiser
📄 [Skill] Vision-in-the-loop LaTeX typesetting agent — auto-compile, render, diagnose, and fix paper layouts