MinerU
opendatalab
Transforms complex documents like PDFs and Office docs into LLM-ready markdown/JSON for your Agentic workflows.
OPENAGENTSKILL / DIRECTORY
Find a skill for your next task. Explore tools for Codex, Claude Code, Cursor and more.
97–112 / 116
Candidates in this shortlist, not the full registry. GitHub stars belong to repositories, not individual skills.
Results: 116
opendatalab
Transforms complex documents like PDFs and Office docs into LLM-ready markdown/JSON for your Agentic workflows.
thanhkeke97
🎮 Real-time Game Translation Tool | OCR + AI Translation | Windows Gaming | Open Source
AkagawaTsurunaki
AI VTuber with LLM, ASR, TTS, OCR, CV and more technologies to live stream or play Minecraft with you.
XMuli
A simple and beautiful cross-platform screenshot software, It also supports OCR, image translation, stickers and pinning images features. | 简单且漂亮的跨平台截图软件,支持离线 OCR、图片翻译、贴…
rtr46
meikipop - universal japanese ocr popup dictionary for windows, linux and macos
run-llama
A fast, helpful, and open-source document parser
pymupdf
PyMuPDF is a high performance Python library for data extraction, analysis, conversion & manipulation of PDF (and other) documents.
ExtractPDF4J
Java PDF table extraction & OCR library. Extract structured tables from text-based and scanned PDFs using stream, lattice (OpenCV-style grid detection), and hybrid parsi…
anig1scur
Add PDF bookmarks online and make it searchable by OCR
bytedance
The official repo for “Dolphin: Document Image Parsing via Heterogeneous Anchor Prompting”, ACL, 2025.
oomol-lab
PDF craft can convert PDF files into various other formats. This project will focus on processing PDF files of scanned books.
getomni-ai
OCR & Document Extraction using vision models
axa-group
Transforms PDF, Documents and Images into Enriched Structured Data
clovaai
Official Implementation of OCR-free Document Understanding Transformer (Donut) and Synthetic Document Generator (SynthDoG), ECCV 2022
cyanfish
Scan documents to PDF and more, as simply as possible.