Extract
ICIJ
A cross-platform command line tool for parallelised content extraction and analysis.
OPENAGENTSKILL / DIRECTORY
Find a skill for your next task. Explore tools for Codex, Claude Code, Cursor and more.
Find a skill for your next task. Preview examples where available.
1–16 / 22
Results: 22
ICIJ
A cross-platform command line tool for parallelised content extraction and analysis.
yifanfeng97
Transform unstructured text into structured knowledge with LLMs. Graphs, hypergraphs, and spatio-temporal extractions — with one command.
CatchTheTornado
Document (PDF, Word, PPTX ...) extraction and parse API using state of the art modern OCRs + Ollama supported models. Anonymize documents. Remove PII. Convert any docume…
extractus
To extract main article from given URL with Node.js
omkarcloud
Google Maps Scraper & Lead Generation Tool. Extract 50+ data points including business emails, phone numbers, and social profiles. Includes enrichment features, API acce…
jsvine
Plumb a PDF for detailed information about each char, rectangle, line, et cetera — and easily extract text and tables.
torakiki
PDFsam, a desktop application to split, merge, mix, rotate PDF files and extract pages
extract internal monitoring data from application logs for collection in a timeseries database
UglyToad
Read and extract text and other content from PDFs in C# (port of PDFBox)
apify
Crawlee—A web scraping and browser automation library for Node.js to build reliable crawlers. In JavaScript and TypeScript. Extract data for AI, LLMs, RAG, or GPTs. Down…
camelot-dev
A web interface to extract tabular data from PDFs
spatie
NanoNets
Extract and convert data from any document, images, pdfs, word doc, ppt or URL into multiple formats (Markdown, JSON, CSV, HTML) with intelligent structured data extract…
apify
Crawlee—A web scraping and browser automation library for Python to build reliable crawlers. Extract data for AI, LLMs, RAG, or GPTs. Download HTML, PDF, JPG, PNG, and o…
oxylabs
Crawl a website starting from a URL, find relevant pages, and extract data – all guided by your natural language prompt.
microlinkhq
The headless Chrome/Chromium driver on top of Puppeteer. Take screenshots, generate PDFs, extract text and HTML with a production-ready API.