Crawlee Python
apify
Crawlee—A web scraping and browser automation library for Python to build reliable crawlers. Extract data for AI, LLMs, RAG, or GPTs. Download HTML, PDF, JPG, PNG, and o…
OPENAGENTSKILL / DIRECTORY
Find a skill for your next task. Explore tools for Codex, Claude Code, Cursor and more.
Find a skill for your next task. Preview examples where available.
1–16 / 19
Results: 19
apify
Crawlee—A web scraping and browser automation library for Python to build reliable crawlers. Extract data for AI, LLMs, RAG, or GPTs. Download HTML, PDF, JPG, PNG, and o…
hect0x7
Python API for JMComic | 提供Python API访问禁漫天堂,同时支持网页端和移动端 | 禁漫天堂GitHub Actions下载器🚀
scrapy
Scrapy, a fast high-level web crawling & scraping framework for Python.
codelucas
newspaper3k is a news, full-text, and article metadata extraction in Python 3. Advanced docs:
unclecode
Web crawling built for AI
jhao104
Python ProxyPool for web spider
alirezamika
A Smart, Automatic, Fast and Lightweight Web Scraper for Python
mherrmann
Lighter web automation with Python
adbar
Python & Command-line tool to gather text and metadata on the Web: Crawling, scraping, extraction, output as CSV, JSON, HTML, MD, TXT, XML
ScrapeGraphAI
getmaxun
🔥 The open-source no-code platform for web scraping, crawling, search and AI data extraction • Turn websites into structured APIs in minutes 🔥
The process of extracting product data from Amazon using Python, including titles, ratings, prices, images, and descriptions.
pymupdf
PyMuPDF is a high performance Python library for data extraction, analysis, conversion & manipulation of PDF (and other) documents.
lexiforest
Python binding for curl-impersonate fork via cffi. A http client that can impersonate browser tls/ja3/http2 fingerprints.
microsoft
Python tool for converting files and office documents to Markdown.
openai
Automate a real browser from the terminal for navigation, form filling, snapshots, screenshots, extraction, and UI-flow debugging.