No visual example yet
Explore the skillDsh Vision Toolkit
Anionex
让纯文本模型更好地做视觉任务的DeepSeek Harness插件:带意图的图片问答、长截图 OCR、UI 还原等|DeepSeek Harness-native integration for agent-vision-toolkit: image Q&A, long-screenshot OCR, UI restoration, g…
OPENAGENTSKILL / DIRECTORY
Find a skill for your next task. Explore tools for Codex, Claude Code, Cursor and more.
9 Skills
Results: 9
No visual example yet
Explore the skillAnionex
让纯文本模型更好地做视觉任务的DeepSeek Harness插件:带意图的图片问答、长截图 OCR、UI 还原等|DeepSeek Harness-native integration for agent-vision-toolkit: image Q&A, long-screenshot OCR, UI restoration, g…
No visual example yet
Explore the skillAnionex
给纯文本 LLM agent 装上眼睛:图片问答、OCR、截图分析、视觉定位等一套视觉工具箱 + skill,并可无缝接入 Codex、Claude Code、OpenCode、Pi | Give text-only LLM agents vision: image Q&A, OCR, screenshot understanding,…
No visual example yet
Explore the skillriddleling
An iOS OCR Server Using Apple’s Vision Framework
No visual example yet
Explore the skillAgents365-ai
Generate draw.io diagrams from natural language — 6 presets, vision self-check + up to 5-round refinement, codebase-to-diagram, 10,000+ official shapes & 321 AI/LLM bran…
No visual example yet
Explore the skillgetomni-ai
OCR & Document Extraction using vision models
No visual example yet
Explore the skillA9T9
Ui.Vision Open-Source RPA Software with Computer Vision, OCR, Anthropic Computer Use/LLM. Selenium IDE import/export.
No visual example yet
Explore the skillicereed
Use LLMs and LLM Vision (OCR) to handle paperless-ngx - Document Digitalization powered by AI
No visual example yet
Explore the skillARahim3
Fine-tune LLMs on your Mac with Apple Silicon. SFT, DPO, GRPO, Vision, TTS, STT, Embedding, and OCR fine-tuning — natively on MLX. Unsloth-compatible API.
No visual example yet
Explore the skillemcf
Get clean data from tricky documents, powered by vision-language models ⚡