No visual example yet
Explore the skillAwesome Ocr
zacharywhitley
zacharywhitley/awesome-ocr is a high-star GitHub project relevant to AI agent workflows.
OPENAGENTSKILL / DIRECTORY
Find a skill for your next task. Explore tools for Codex, Claude Code, Cursor and more.
Find a skill for your next task. Preview examples where available.
209–224 / 481
Results: 481
No visual example yet
Explore the skillzacharywhitley
zacharywhitley/awesome-ocr is a high-star GitHub project relevant to AI agent workflows.
No visual example yet
Explore the skillborb-pdf
borb is a library for reading, creating and manipulating PDF files in python.
No visual example yet
Explore the skillAgents365-ai
Mermaid diagrams (.mmd) from natural language with validation loop. 11+ types, multi-backend (mmdc / Kroki), PNG/SVG/PDF, multi-agent.
No visual example yet
Explore the skillDicklesworthstone
Enhances Tesseract OCR output using LLMs (local or API) for error correction, smart chunking, and markdown formatting of scanned PDFs
No visual example yet
Explore the skillMuiseDestiny
No visual example yet
Explore the skillzubair-trabzada
AI Marketing Suite for Claude Code. 15 marketing skills with parallel subagents — audit any website, generate copy, email sequences, ad campaigns, content calendars, com…
No visual example yet
Explore the skilldynobo
OCR powered screen-capture tool to capture information instead of images
No visual example yet
Explore the skillripperhe
Bob 是一款 macOS 平台的翻译和 OCR 软件。
No visual example yet
Explore the skillszTheory
Cross-platform desktop GUI app to clean image metadata
No visual example yet
Explore the skilleikek
Assist in organizing your piles of documents, resulting from scanners, e-mails and other sources with miminal effort.
No visual example yet
Explore the skillmodesty
converts binary PDF to JSON and text, for server-side PDF processing and command-line use. Zero dependency.
No visual example yet
Explore the skillsirfz
A Python wrapper for the tesseract-ocr API
No visual example yet
Explore the skillTimmyOVO
Rust multi‑backend OCR/VLM engine (DeepSeek‑OCR-1/2, PaddleOCR‑VL, DotsOCR) with DSQ quantization and an OpenAI‑compatible server & CLI – run locally without Python.
No visual example yet
Explore the skilladithya-s-k
Ingest, parse, and optimize any data format ➡️ from documents to multimedia ➡️ for enhanced compatibility with GenAI frameworks
No visual example yet
Explore the skillNanoNets
An on-premises, OCR-free unstructured data extraction, markdown conversion and benchmarking toolkit. (https://idp-leaderboard.org/)
No visual example yet
Explore the skilllukas-blecher
pix2tex: Using a ViT to convert images of equations into LaTeX code.
Design direction, Figma implementation, UI review, React performance, browser QA, and safe preview deployments.
Code review, repo inspection, testing, planning, shipping, and engineering workflows for coding agents.
Research, recent-events briefings, PDF parsing, markdown conversion, RAG ingestion, and knowledge workflows.
Crawling, scraping, extraction, browser automation, structured data capture, and website-to-markdown workflows.
Presentation generation, editable PPTX, slide decks, speaker notes, and visual storytelling workflows.
Choose React motion graphics, generated footage or editing; compare dependencies and inspect actual outputs.
Image, video, creative production, UI design, multimodal generation, and visual workflow skills.
Data analysis, analytics, ETL, notebooks, databases, tables, charts, and reporting workflows.