PaddleOCR
PaddlePaddle
Turn any PDF or image document into structured data for your AI. A powerful, lightweight OCR toolkit that bridges the gap between images/PDFs and LLMs. Supports 100+ lan…
OPENAGENTSKILL / DIRECTORY
Find a skill for your next task. Explore tools for Codex, Claude Code, Cursor and more.
Find a skill for your next task. Preview examples where available.
1–11 / 11
Results: 11
PaddlePaddle
Turn any PDF or image document into structured data for your AI. A powerful, lightweight OCR toolkit that bridges the gap between images/PDFs and LLMs. Supports 100+ lan…
microsoft
Python tool for converting files and office documents to Markdown.
firecrawl
The API to search, scrape, and interact with the web at scale. 🔥
iamgio
🪐 Markdown with superpowers: from ideas to papers, presentations, websites, books, and knowledge bases.
Stirling-Tools
#1 PDF Application on GitHub that lets you edit PDFs on any device anywhere
tesseract-ocr
Tesseract Open Source OCR Engine (main repository)
opendatalab
Transforms complex documents like PDFs and Office docs into LLM-ready markdown/JSON for your Agentic workflows.
apify
Crawlee—A web scraping and browser automation library for Node.js to build reliable crawlers. In JavaScript and TypeScript. Extract data for AI, LLMs, RAG, or GPTs. Down…
apify
Crawlee—A web scraping and browser automation library for Python to build reliable crawlers. Extract data for AI, LLMs, RAG, or GPTs. Download HTML, PDF, JPG, PNG, and o…
getmaxun
🔥 The open-source no-code platform for web scraping, crawling, search and AI data extraction • Turn websites into structured APIs in minutes 🔥
unclecode
Open-source LLM-friendly web crawler and scraper for agent workflows.