暂未收录效果图
查看技能说明Dsh Vision Toolkit
Anionex
让纯文本模型更好地做视觉任务的DeepSeek Harness插件:带意图的图片问答、长截图 OCR、UI 还原等|DeepSeek Harness-native integration for agent-vision-toolkit: image Q&A, long-screenshot OCR, UI restoration, g…
OPENAGENTSKILL / DIRECTORY
为下一项任务找到合适的技能。探索适用于 Codex、Claude Code、Cursor 等 Agent 的工具。
9 Skills
搜索结果: 9
暂未收录效果图
查看技能说明Anionex
让纯文本模型更好地做视觉任务的DeepSeek Harness插件:带意图的图片问答、长截图 OCR、UI 还原等|DeepSeek Harness-native integration for agent-vision-toolkit: image Q&A, long-screenshot OCR, UI restoration, g…
暂未收录效果图
查看技能说明clovaai
Official Implementation of OCR-free Document Understanding Transformer (Donut) and Synthetic Document Generator (SynthDoG), ECCV 2022
暂未收录效果图
查看技能说明riddleling
暂未收录效果图
查看技能说明getomni-ai
暂未收录效果图
查看技能说明icereed
Use LLMs and LLM Vision (OCR) to handle paperless-ngx - Document Digitalization powered by AI
暂未收录效果图
查看技能说明emcf
Get clean data from tricky documents, powered by vision-language models ⚡
暂未收录效果图
查看技能说明czlonkowski
Handle files and binary data in n8n correctly. Use when working with files, images, PDFs, attachments, uploads or downloads, base64, vision/multimodal input, or when an…
暂未收录效果图
查看技能说明jamjamjon
A Rust library integrated with ONNXRuntime, providing a collection of Computer Vison and Vision-Language models such as YOLO, FastVLM, and more.
暂未收录效果图
查看技能说明qpdf