Open-source LLM-friendly web crawler and scraper
Skill 디렉토리
AI Agent를 위한 재사용 가능한 Skill을 찾으세요.
모든 추천은 리포지토리, 감사, 설치 경로와 명확하게 연결됩니다.
검색 결과: structured-ocr
영문 디렉토리Review a branch or diff against repository standards and the originating spec in two independent analysis passes.
Turn the current conversation and codebase context into a structured implementation spec, then publish it to the configured project issue tracker.
A versatile command-line tool for interacting with Google Workspace APIs, designed for both human users and AI agents.
Turn any PDF or image document into structured data for your AI. A powerful, lightweight OCR toolkit that bridges the gap between images/PDFs and LLMs. Supports 100+ languages.
Tesseract Open Source OCR Engine (main repository)
LlamaIndex is the leading document agent and OCR platform
Official Lark/Feishu CLI tool with 200+ commands and 26 AI agent skills, designed for agent-native operation and easy integration with AI runtimes.
Weaviate is an open-source vector database that stores both objects and vectors, allowing for the combination of vector search with structured filtering with the fault tolerance and scalability of a cloud-native database.
Turn any website into LLM-ready markdown or structured data
OCRmyPDF adds an OCR text layer to scanned PDF files, allowing them to be searched
A self-contained AI agent skill for accessing China A-share stock data from 15 sources, packaged as a structured Markdown+Python file for Claude Code, Codex, and similar agent runtimes.