PDF to markdown using vision LLMs — tables, layouts, and structure preserved
Supply asset profile
Deep research, source comparison, literature review, RAG, knowledge search, and reports.
Scenario
Document processing
I need my agent to read PDFs, extract tables, and turn documents into structured data.
Agent fit
Claude Code + CLI + Codex
Codex, Claude Code, Cursor, CLI, or custom agents.
Install
Ready
npx skills add yigitkonur/api-llm-ocr
Maintenance
active
5mo since push
Risk
Needs review
License is unclear
GitHub quality
895
72/100 quality · 74/100 trust
Coverage tags
Review notes
License is unclear · Quality score needs review
Agent adoption scorecard
These scores combine public repository metadata, OpenAgentSkill review signals, maintenance freshness, and install readiness. They are a shortlist signal, not a replacement for human review.
Quality
StrongSolid option that is likely worth shortlisting for production workflows.
Trust
Sandbox onlyUseful candidate with missing or mixed trust signals. Keep it in an isolated workspace until the outcome loop proves task fit.
Audit
Needs reviewInstall readiness, security metadata, maintenance, and adoption risk.
Trust Score v5
Run only in a sandbox and compare close alternatives before using it for real work.
Stars
895 GitHub stars
Repo activity
895 stars, 60 forks
Maintenance
5mo since push
License
Unknown
Install
npx skills add yigitkonur/api-llm-ocr
Install safety
standard package or runtime install path
Permission surface
filesystem or document access, network or browser access
Agent outcomes
No agent outcome data yet
Docs
Usable metadata, review docs
Risk summary
Install readiness
Agent-readable metadata
Use this block or the embedded JSON to decide whether an agent should install this skill, choose an alternative, or ask for human review first.
Suited tasks
Suited agents
Install decision
Trust and risk
Outcome loop
Install command
npx skills add yigitkonur/api-llm-ocrDo not use when
Agent safety v2
Usable candidate, but the agent should surface permission and audit notes before installation.
Require human approval before installing into a real workspace.
medium
Skill likely fetches remote pages, APIs, repositories, or external services.
medium
Skill may read or write project files, documents, generated artifacts, or local workspace state.
Install targets
Copy the registry command or an agent-specific install prompt for Codex, Claude Code, and Cursor.
Use the registry command when your workflow supports the OpenAgentSkill installer.
$ npx skills add yigitkonur/api-llm-ocrAgent resolve plan
The Resolve API returns the selected skill, alternatives, safety policy, audit notes, install target, and copy-paste prompt an agent can follow without scraping this page.
Resolve JSON
/api/agent/resolve?task=Use%20API%20Llm%20Ocr%20for%20an%20agent%20workflow&agent=codex&max_risk=medium
Resolve text
/api/agent/resolve?task=Use%20API%20Llm%20Ocr%20for%20an%20agent%20workflow&agent=codex&max_risk=medium&format=text
Install handoff
/api/skills/yigitkonur-api-llm-ocr/install
Agent should check
Copy prompt
Task: Use API Llm Ocr in this workspace.
Resolve first: https://www.openagentskill.com/api/agent/resolve?task=Use%20API%20Llm%20Ocr%20for%20an%20agent%20workflow&agent=codex&max_risk=medium
Review install handoff: https://www.openagentskill.com/api/skills/yigitkonur-api-llm-ocr/install
Install command: npx skills add yigitkonur/api-llm-ocr
Before running it, summarize audit warnings, required permissions, and the fallback skill if install is risky.Agent handoff
Use the public install endpoint to fetch the command, safety checklist, target prompts, and canonical links for this skill.
Install handoff
/api/skills/yigitkonur-api-llm-ocr/install
LLM text format
/api/skills/yigitkonur-api-llm-ocr/install?format=text
Find alternatives
/api/skills/search?q=API%20Llm%20Ocr&limit=3
Agent prompt
Use API Llm Ocr for this task. Review https://www.openagentskill.com/api/skills/yigitkonur-api-llm-ocr/install, then install with: npx skills add yigitkonur/api-llm-ocrRegistry metadata
This page exposes the same decision, trust, audit, use-case, and install signals through the Registry API, so agents can rank this skill without scraping the UI.
Manifest
/api/registry/manifest/yigitkonur-api-llm-ocr
LLM text
/api/registry/manifest/yigitkonur-api-llm-ocr?format=text
Install alias
/api/registry/install/yigitkonur-api-llm-ocr
Recommend
/api/registry/recommend?task=Use%20API%20Llm%20Ocr%20in%20an%20agent%20workflow&limit=3
Agent fit
Document processing
Use-case tags
Platforms
Python, OCR, Claude Code
Audit report
Review install readiness, maintenance, trust, quality, and metadata warnings before adding this skill to an agent workflow.
Agent decision cockpit
Use this as a leading candidate, then validate the README and install path in your own agent stack.
Role in stack
Primary pick
Primary fit
Document processing
Trust label
Production-ready
Install path
Command ready
Use when
Evidence
Review first
Implementation path
Trust profile
Useful candidate with missing or mixed trust signals. Keep it in an isolated workspace until the outcome loop proves task fit.
GitHub adoption
INFO895 GitHub stars
Stars/forks activity
INFO895 stars, 60 forks; issue activity unavailable in current metadata
Recent maintenance
INFO5mo since push
License clarity
CHECKUnknown
Good signals
Review before install
Recommended action
Run only in a sandbox and compare close alternatives before using it for real work.
Quality profile
Solid option that is likely worth shortlisting for production workflows.
Workflow fit
Parse messy files
I need my agent to read PDFs, extract tables, and turn documents into structured data.
Search private knowledge
I need my agent to build a RAG workflow over documents and retrieve reliable context.
Build and ship code
I need a coding agent that can understand a repository, edit code, and review pull requests.
Stack fit
Ingest, retrieve, and cite
A stack for document-heavy agents that ingest files, create searchable knowledge, retrieve relevant context, and answer with grounded sources.
Turn skills into distribution
A stack for turning newly indexed skills into SEO briefs, social drafts, comparison pages, and reusable publishing workflows.
Scrape, clean, and reuse web data
A practical stack for agents that crawl public pages, extract clean content, normalize data, and hand it to downstream research or RAG workflows.
Alternative shortlist
Similar skills in this category, ranked with the same readiness and quality signals.
Python tool for converting files and office documents to Markdown.
Turn any PDF or image document into structured data for your AI. A powerful, lightweight OCR toolkit that bridges the gap between images/PDFs and LLMs. Supports 100+ languages.
#1 PDF Application on GitHub that lets you edit PDFs on any device anywhere
Tesseract Open Source OCR Engine (main repository)
PDF to markdown using vision LLMs — tables, layouts, and structure preserved
Imported by the skill-only GitHub discovery pipeline because it matches agent skill, automation, domain workflow, RAG, document-processing, data, finance, security, or developer-tool signals. Protocol-server projects are excluded from automated imports.
Frameworks & Tools
Decision snapshot
895 GitHub stars
Audit snapshot
Install and adoption review
Agent-proven evidence
Outcome reports after resolve, review, install, and one narrow run.
No agent outcome data yet. The first agent run can report success, setup needs, risk blocks, failure, or not-relevant through /api/agent/outcome.
Install
Free and open source. Review the audit before production use.
Growth loop
Scenario-led draft for API Llm Ocr, ready for a manual X post.
The useful research skills are not search wrappers. They help agents keep sources attached. API Llm Ocr helps agents turn docs, data, or knowledge bases into grounded work. 895 stars https://www.openagentskill.com/skills/yigitkonur-api-llm-ocr?ref=x #AIAgents
Listing + install path for API Llm Ocr: https://www.openagentskill.com/skills/yigitkonur-api-llm-ocr?ref=x Install: npx skills add yigitkonur/api-llm-ocr
Listing source
This listing was indexed from public sources and is not marked official until a maintainer claim is approved.
Attribution links to the public repository or creator profile. Creators can claim the listing to update ownership signals.
Claim this skillOwner claim
This community indexed listing is attributed to yigitkonur but is not marked official yet. Claim it to add a verified owner signal and make future launch, install, and audit updates easier to trust.
Creator backlink kit
Show the canonical listing, current trust and audit signals, and real Agent Proven evidence where developers evaluate the repository.
[](https://www.openagentskill.com/skills/yigitkonur-api-llm-ocr)
[](https://www.openagentskill.com/skills/yigitkonur-api-llm-ocr)
[](https://www.openagentskill.com/skills/yigitkonur-api-llm-ocr/audit)
[](https://www.openagentskill.com/skills/yigitkonur-api-llm-ocr)yigitkonur
@yigitkonur
Share whether this skill looks useful for your agent workflow. Aggregated feedback improves rankings over time.
Sandbox only
Markitdown
Python tool for converting files and office documents to Markdown.
156.1K stars · 0 installsPaddleOCR
Turn any PDF or image document into structured data for your AI. A powerful, lightweight OCR toolkit that bridges the gap between images/PDFs and LLMs. Supports 100+ languages.
83.1K stars · 0 installsStirling PDF
#1 PDF Application on GitHub that lets you edit PDFs on any device anywhere
81.2K stars · 0 installsTesseract
Tesseract Open Source OCR Engine (main repository)
74.7K stars · 0 installs