Trafilatura
adbar
Python & Command-line tool to gather text and metadata on the Web: Crawling, scraping, extraction, output as CSV, JSON, HTML, MD, TXT, XML
OPENAGENTSKILL / DIRECTORY
Find a skill for your next task. Explore tools for Codex, Claude Code, Cursor and more.
Find a skill for your next task. Preview examples where available.
1–14 / 14
Results: 14
adbar
Python & Command-line tool to gather text and metadata on the Web: Crawling, scraping, extraction, output as CSV, JSON, HTML, MD, TXT, XML
unclecode
Open-source LLM-friendly web crawler and scraper for agent workflows.
mendableai
Turn websites into clean markdown or structured data for retrieval and agents.
apify
Crawlee—A web scraping and browser automation library for Node.js to build reliable crawlers. In JavaScript and TypeScript. Extract data for AI, LLMs, RAG, or GPTs. Down…
apify
Crawlee—A web scraping and browser automation library for Python to build reliable crawlers. Extract data for AI, LLMs, RAG, or GPTs. Download HTML, PDF, JPG, PNG, and o…
openai
Automate a real browser from the terminal for navigation, form filling, snapshots, screenshots, extraction, and UI-flow debugging.
microlinkhq
The headless Chrome/Chromium driver on top of Puppeteer. Take screenshots, generate PDFs, extract text and HTML with a production-ready API.
pinchtab
High-performance browser automation bridge and multi-instance orchestrator with advanced stealth injection and real-time dashboard.
mpdf
PHP library generating PDF files from UTF-8 encoded HTML
astefanutti
lmn1919
Convert HTML to a multi-thousand-page vector PDF with a single line of frontend code
plutoprint
A Python Library for Generating PDFs and Images from HTML, powered by PlutoBook
wechat-article
微信公众号文章批量下载工具,支持导出阅读量与评论数据。无需搭建环境,支持在线使用、Docker 私有化部署和 Cloudflare 部署。支持多种格式导出,HTML 格式可100%还原文章排版与样式。
LibrePDF
OpenPDF is an open-source Java library for creating, editing, rendering, and encrypting PDF documents, as well as generating PDFs from HTML. It is licensed under the LGP…