Crawlee
apify
Crawlee—A web scraping and browser automation library for Node.js to build reliable crawlers. In JavaScript and TypeScript. Extract data for AI, LLMs, RAG, or GPTs. Down…
OPENAGENTSKILL / DIRECTORY
Find a skill for your next task. Explore tools for Codex, Claude Code, Cursor and more.
Find a skill for your next task. Preview examples where available.
Page 1 · 16 shown · 1,600 public entries
Results: 1600
apify
Crawlee—A web scraping and browser automation library for Node.js to build reliable crawlers. In JavaScript and TypeScript. Extract data for AI, LLMs, RAG, or GPTs. Down…
firecrawl
The API to search, scrape, and interact with the web at scale. 🔥
PaddlePaddle
Turn any PDF or image document into structured data for your AI. A powerful, lightweight OCR toolkit that bridges the gap between images/PDFs and LLMs. Supports 100+ lan…
browser-use
Give your AI agent a web browser
unclecode
Web crawling built for AI
ScrapeGraphAI
microsoft
Python tool for converting files and office documents to Markdown.
getmaxun
🔥 The open-source no-code platform for web scraping, crawling, search and AI data extraction • Turn websites into structured APIs in minutes 🔥
Stirling-Tools
#1 PDF Application on GitHub that lets you edit PDFs on any device anywhere
apify
Crawlee—A web scraping and browser automation library for Python to build reliable crawlers. Extract data for AI, LLMs, RAG, or GPTs. Download HTML, PDF, JPG, PNG, and o…
iamgio
🪐 Markdown with superpowers: from ideas to papers, presentations, websites, books, and knowledge bases.
D4Vinci
Adaptive web scraping for agent data collection
openai
Automate a real browser from the terminal for navigation, form filling, snapshots, screenshots, extraction, and UI-flow debugging.
tesseract-ocr
Tesseract Open Source OCR Engine (main repository)
opendatalab
Transforms complex documents like PDFs and Office docs into LLM-ready markdown/JSON for your Agentic workflows.
docling-project
Get your documents ready for gen AI