Open-source LLM-friendly web crawler and scraper
Skill-Verzeichnis
Wiederverwendbare Skills für AI Agents entdecken.
Jede Empfehlung bleibt mit ihrem Repository, Audit und Installationspfad nachvollziehbar.
Suchergebnisse: structured-ocr
Englisches VerzeichnisReview a branch or diff against repository standards and the originating spec in two independent analysis passes.
Turn the current conversation and codebase context into a structured implementation spec, then publish it to the configured project issue tracker.
A versatile command-line tool for interacting with Google Workspace APIs, designed for both human users and AI agents.
Turn any PDF or image document into structured data for your AI. A powerful, lightweight OCR toolkit that bridges the gap between images/PDFs and LLMs. Supports 100+ languages.
Tesseract Open Source OCR Engine (main repository)
LlamaIndex is the leading document agent and OCR platform
Official Lark/Feishu CLI tool with 200+ commands and 26 AI agent skills, designed for agent-native operation and easy integration with AI runtimes.
Weaviate is an open-source vector database that stores both objects and vectors, allowing for the combination of vector search with structured filtering with the fault tolerance and scalability of a cloud-native database.
Turn any website into LLM-ready markdown or structured data
OCRmyPDF adds an OCR text layer to scanned PDF files, allowing them to be searched
A self-contained AI agent skill for accessing China A-share stock data from 15 sources, packaged as a structured Markdown+Python file for Claude Code, Codex, and similar agent runtimes.