News Please
fhamborg
news-please - an integrated web crawler and information extractor for news that just works
OPENAGENTSKILL / DIRECTORY
Find a skill for your next task. Explore tools for Codex, Claude Code, Cursor and more.
Find a skill for your next task. Preview examples where available.
1–16 / 57
Results: 57
fhamborg
news-please - an integrated web crawler and information extractor for news that just works
codelucas
newspaper3k is a news, full-text, and article metadata extraction in Python 3. Advanced docs:
bda-research
Web Crawler/Spider for NodeJS + server-side jQuery ;-)
AndyTheFactory
📰 Newspaper4k a fork of the beloved Newspaper3k. Extraction of articles, titles, and metadata from news websites.
polyrabbit
:newspaper: Let ChatGPT Summarize Hacker News for You
hiddendevj
Collection of China illegal cases about web crawler 本项目用来整理所有中国大陆爬虫开发者涉诉与违规相关的新闻、资料与法律法规。致力于帮助在中国大陆工作的爬虫行业从业者了解我国相关法律,避免触碰数据合规红线。
BruceDone
A collection of awesome web crawler,spider in different languages
YoongiKim
Google, Naver multiprocess image web crawler (Selenium)
apache
A scalable, mature and versatile web crawler based on Apache Storm
fredwu
A high performance web crawler / scraper in Elixir.
xuxueli
A lightweight web crawler framework.(Java爬虫框架)
hect0x7
Python API for JMComic | 提供Python API访问禁漫天堂,同时支持网页端和移动端 | 禁漫天堂GitHub Actions下载器🚀
dataabc
新浪微博爬虫,用python爬取新浪微博数据,并下载微博图片和微博视频
kanasimi
Download comics novels 小说漫画下载工具 小説漫画のダウンローダ 小說漫畫下載:腾讯漫画 大角虫漫画 有妖气 咪咕 SF漫画 哦漫画 看漫画 漫画柜 汗汗酷漫 動漫伊甸園 快看漫画 微博动漫 733动漫网 大古漫画网 漫画DB 無限動漫 動漫狂 卡推漫画 动漫之家 动漫屋 古风漫画网 36漫画网 亲亲漫画网 乙女漫…
JayBizzle
🕷 CrawlerDetect is a PHP class for detecting bots/crawlers/spiders via the user agent
dadoonet
Elasticsearch File System Crawler (FS Crawler)