Metadata Action
docker
GitHub Action to extract metadata (tags, labels) from Git reference and GitHub events for Docker
OPENAGENTSKILL / DIRECTORY
Find a skill for your next task. Explore tools for Codex, Claude Code, Cursor and more.
Search retrieves candidates across the registry and ranks a bounded shortlist by task fit. This count is matching candidates, not the registry total. No suitable match? Try a specific tool or task.
1–16 / 38
Results: 38
docker
GitHub Action to extract metadata (tags, labels) from Git reference and GitHub events for Docker
pymupdf
PyMuPDF is a high performance Python library for data extraction, analysis, conversion & manipulation of PDF (and other) documents.
firecrawl
Fast Rust library for PDF inspection, classification, and text extraction. Intelligently detects scanned vs text-based PDFs to enable smart routing decisions.
adbar
Python & Command-line tool to gather text and metadata on the Web: Crawling, scraping, extraction, output as CSV, JSON, HTML, MD, TXT, XML
landing-ai
This tool has been deprecated. Use Agentic Document Extraction instead.
JimmySadek
Claude Code skill: turn YouTube videos into structured, Obsidian-ready Markdown notes with full metadata, chapters, and transcripts
codelucas
newspaper3k is a news, full-text, and article metadata extraction in Python 3. Advanced docs:
windingwind
Translate PDF, EPub, webpage, metadata, annotations, notes to the target language. Support 20+ translate services.
szTheory
Cross-platform desktop GUI app to clean image metadata
AndyTheFactory
📰 Newspaper4k a fork of the beloved Newspaper3k. Extraction of articles, titles, and metadata from news websites.
Achno
A tool to convert a Wallpaper's color scheme / palette, OCR with VLM's Traditional & Hybrid, Image Compression ,color palette extraction, image upsacling with Adversaria…
NanoNets
An on-premises, OCR-free unstructured data extraction, markdown conversion and benchmarking toolkit. (https://idp-leaderboard.org/)
Querybook is a Big Data Querying UI, combining collocated table metadata and a simple notebook interface.
landing-ai
Python library for Agentic Document Extraction (ADE).
getmaxun
🔥 The open-source no-code platform for web scraping, crawling, search and AI data extraction • Turn websites into structured APIs in minutes 🔥
Owner-curated external sources. Not filtered by the scores or compatibility controls above; excluded from GitHub rankings and automatic installation.
No external entries match this query.