Browsertrix Crawler
webrecorder
Run a high-fidelity browser-based web archiving crawler in a single Docker container
OPENAGENTSKILL / DIRECTORY
Find a skill for your next task. Explore tools for Codex, Claude Code, Cursor and more.
Find a skill for your next task. Preview examples where available.
17–32 / 85
Results: 85
webrecorder
Run a high-fidelity browser-based web archiving crawler in a single Docker container
Qianlitp
A powerful browser crawler for web vulnerability scanners
scrapy
Scrapy, a fast high-level web crawling & scraping framework for Python.
dixudx
Easily download all the photos/videos from tumblr blogs. 下载指定的 Tumblr 博客中的图片,视频
alecxe
Random User-Agent middleware based on fake-useragent
paulpierre
A multithreaded 🕸️ web crawler that recursively crawls a website and creates a 🔽 markdown file for each page, designed for LLM RAG
janreges
SiteOne Crawler is a cross-platform website crawler and analyzer for SEO, security, accessibility, and performance optimization—ideal for developers, DevOps, QA engineer…
NanmiCoder
小红书笔记 | 评论爬虫、抖音视频 | 评论爬虫、快手视频 | 评论爬虫、B站视频 | 评论爬虫、微博帖子 | 评论爬虫、百度贴吧帖子 | 知乎问答文章 | 评论爬虫。支持多平台社交媒体内容抓取,提供完整的数据采集解决方案。
wycm
zhihu-crawler是一个基于Java的高性能、支持免费http代理池、支持横向扩展、分布式爬虫项目
oxylabs
Crawl a website starting from a URL, find relevant pages, and extract data – all guided by your natural language prompt.
yujiosaka
Distributed crawler powered by Headless Chrome
Boris-code
🚀🚀🚀feapder is an easy to use, powerful crawler framework | feapder是一款上手简单,功能强大的Python爬虫框架。内置AirSpider、Spider、TaskSpider、BatchSpider四种爬虫解决不同场景的需求。且支持断点续爬、监控报警、浏览器渲染、海量…
SpiderClub
:sparkling_heart: High available distributed ip proxy pool, powerd by Scrapy and Redis
code4craft
A scalable web crawler framework for Java.
fhamborg
news-please - an integrated web crawler and information extractor for news that just works
xtuhcy
Easy to use lightweight web crawler(易用的轻量化网络爬虫)