OPENAGENTSKILL / DIRECTORY

AI Agent Skills

Encuentra una habilidad para tu próxima tarea con Codex, Claude Code, Cursor y más.

Resultados · “ocr”

101 Skills

Resultados: 101

Documentos

No visual example yet

Explore the skill

让纯文本模型更好地做视觉任务的DeepSeek Harness插件:带意图的图片问答、长截图 OCR、UI 还原等|DeepSeek Harness-native integration for agent-vision-toolkit: image Q&A, long-screenshot OCR, UI restoration, g…

Precio sin confirmarDocumentosClaude CodeRevisar antes de usar
819GitHub
Documentos

No visual example yet

Explore the skill

Texo

alephpi

A minimalist SOTA LaTeX OCR model with only 20M parameters, running in browser. Full training pipeline available for self-reproduction. | 超轻量SOTA LaTeX公式识别模型,仅20M参数量,可在浏…

Precio sin confirmarDocumentosBrowser agentsRevisar antes de usar
883GitHub
Documentos

No visual example yet

Explore the skill

🖼️ Image Toolbox is a powerful app for advanced image manipulation. It offers dozens of features, from basic tools like crop and draw to filters, OCR, and a wide range…

Precio sin confirmarDocumentosRevisar antes de usar
13,2 milGitHub
Productividad

No visual example yet

Explore the skill

Windrecorder

yuka-friends

Windrecorder is a memory search app by records everything on your screen in small size, to let you rewind what you have seen, query through OCR text or image description…

Precio sin confirmarProductividadRevisar antes de usar
3,9 milGitHub
Desarrollo

No visual example yet

Explore the skill

给纯文本 LLM agent 装上眼睛:图片问答、OCR、截图分析、视觉定位等一套视觉工具箱 + skill,并可无缝接入 Codex、Claude Code、OpenCode、Pi | Give text-only LLM agents vision: image Q&A, OCR, screenshot understanding,…

Precio sin confirmarDesarrolloClaude CodeOpenAI AgentsRevisar antes de usar
1,1 milGitHub
Documentos

No visual example yet

Explore the skill

odl-pdf

opendataloader-project

Procedure for extracting structured data from PDFs with opendataloader-pdf (ODL) correctly: read the installed tool's own help to discover its current options, build the…

Precio sin confirmarDocumentosClaude Code
28,9 milGitHub
Documentos

No visual example yet

Explore the skill

pdf

eigent-ai

Use this skill whenever the user wants to do anything with PDF files. This includes reading or extracting text/tables from PDFs, combining or merging multiple PDFs into…

Precio sin confirmarDocumentosClaude CodeRevisar antes de usar
15,2 milGitHub
Documentos

No visual example yet

Explore the skill

hive.pdf

aden-hive

Read, write, merge, split, rotate, watermark, encrypt, and OCR PDF files using Python (pypdf, pdfplumber, reportlab, pypdfium2) and command-line tools (poppler-utils, qp…

Precio sin confirmarDocumentosClaude CodeBloqueada
11 milGitHub
Documentos

No visual example yet

Explore the skill

modlens

liustack

Plug-in vision for text-only models. Hard rule: when a file path or URL with an image extension (.png, .jpg, .jpeg, .webp, .gif, .heic, .heif) appears anywhere in the co…

Precio sin confirmarDocumentosClaude CodeOpenAI AgentsBloqueada
3,3 milGitHub

Guías y comparativas

Para desarrolladores