These pages are built for high-intent search and for agents that need a structured shortlist with install commands, trust signals, audit links, and real outcome evidence before installing third-party code.
01
Scrape public pricing pages into JSON or markdown
02
Crawl documentation sites for RAG ingestion
03
Monitor public web pages and extract changed fields
04
Use browser automation when static HTML is not enough
Crawlee—A web scraping and browser automation library for Node.js to build reliable crawlers. In JavaScript and TypeScript. Extract data for AI, LLMs, RAG, or GPTs. Download HTML, PDF, JPG, PNG, and other files from websites. Works with Puppeteer, Playwright, Cheerio, JSDOM, and raw HTTP. Both headful and headless mode. With proxy rotation.
100
Quality
88
Trust
—
Proven
browser-automationJun 24, 2026 pushApache-2.0Needs first agent run
Crawlee—A web scraping and browser automation library for Python to build reliable crawlers. Extract data for AI, LLMs, RAG, or GPTs. Download HTML, PDF, JPG, PNG, and other files from websites. Works with Parsel, BeautifulSoup, Playwright, and raw HTTP. Both headful and headless mode. With proxy rotation.
100
Quality
86
Trust
—
Proven
browser-automationJun 19, 2026 pushApache-2.0Needs first agent run
web-automationJul 15, 2026 pushApache-2.0Early agent signal
$ npx skills add unclecode/crawl4ai
Evaluation
How to choose the right skill.
Scope controls for crawl depth and allowed domains
Clean markdown or structured output support
Fresh maintenance and visible repository activity
Clear install command and sandbox-friendly usage
Questions
What is the best web scraping skill for an AI agent?
Start with skills that return clean markdown or structured data, expose crawl limits, and have strong maintenance signals. OpenAgentSkill ranks candidates by task fit, trust, quality, stars, and install readiness.
Can an agent use these skills automatically?
Yes. Agents can call the Resolve API with a scraping task, inspect the ranked shortlist, and fetch install handoffs before running anything.
These pages are built for high-intent search and for agents that need a structured shortlist with install commands, trust signals, audit links, and real outcome evidence before installing third-party code.
01
Scrape public pricing pages into JSON or markdown
02
Crawl documentation sites for RAG ingestion
03
Monitor public web pages and extract changed fields
04
Use browser automation when static HTML is not enough
Crawlee—A web scraping and browser automation library for Node.js to build reliable crawlers. In JavaScript and TypeScript. Extract data for AI, LLMs, RAG, or GPTs. Download HTML, PDF, JPG, PNG, and other files from websites. Works with Puppeteer, Playwright, Cheerio, JSDOM, and raw HTTP. Both headful and headless mode. With proxy rotation.
100
Quality
88
Trust
—
Proven
browser-automationJun 24, 2026 pushApache-2.0Needs first agent run
Crawlee—A web scraping and browser automation library for Python to build reliable crawlers. Extract data for AI, LLMs, RAG, or GPTs. Download HTML, PDF, JPG, PNG, and other files from websites. Works with Parsel, BeautifulSoup, Playwright, and raw HTTP. Both headful and headless mode. With proxy rotation.
100
Quality
86
Trust
—
Proven
browser-automationJun 19, 2026 pushApache-2.0Needs first agent run
web-automationJul 15, 2026 pushApache-2.0Early agent signal
$ npx skills add unclecode/crawl4ai
Evaluation
How to choose the right skill.
Scope controls for crawl depth and allowed domains
Clean markdown or structured output support
Fresh maintenance and visible repository activity
Clear install command and sandbox-friendly usage
Questions
What is the best web scraping skill for an AI agent?
Start with skills that return clean markdown or structured data, expose crawl limits, and have strong maintenance signals. OpenAgentSkill ranks candidates by task fit, trust, quality, stars, and install readiness.
Can an agent use these skills automatically?
Yes. Agents can call the Resolve API with a scraping task, inspect the ranked shortlist, and fetch install handoffs before running anything.