Skill-Verzeichnis

Wiederverwendbare Skills für AI Agents entdecken.

Durchsuche reale GitHub-Skills nach Aufgabe und prüfe Stars, Trust, Audit, Kategorie und Installationspfad vor der Verwendung.

Jede Empfehlung bleibt mit ihrem Repository, Audit und Installationspfad nachvollziehbar.

Suchergebnisse: tesseract

Englisches Verzeichnis

ICCV 2025 | TesserAct: Learning 4D Embodied World Models

400
Stars
62/100
Trust
Kategorie: robotics-iotAudit

Motion Planning Environment

367
Stars
65/100
Trust
Kategorie: robotics-iotAudit

Tesseract Open Source OCR Engine (main repository)

75K
Stars
81/100
Trust
Kategorie: document-processingAudit

Pure Javascript OCR for more than 100 Languages 📖🎉🖥

38K
Stars
81/100
Trust
Kategorie: document-processingAudit

Enhances Tesseract OCR output using LLMs (local or API) for error correction, smart chunking, and markdown formatting of scanned PDFs

2.9K
Stars
73/100
Trust
Kategorie: document-processingAudit

A Python wrapper for the tesseract-ocr API

2.2K
Stars
75/100
Trust
Kategorie: document-processingAudit

Go package for OCR (Optical Character Recognition), by using Tesseract C++ library

3.1K
Stars
79/100
Trust
Kategorie: document-processingAudit

Trained models with fast variant of the "best" LSTM models + legacy models

7.6K
Stars
71/100
Trust
Kategorie: document-processingAudit

A Gtk/Qt front-end to tesseract-ocr.

2.0K
Stars
73/100
Trust
Kategorie: document-processingAudit

Plug-in vision for text-only models. Hard rule: when a file path or URL with an image extension (.png, .jpg, .jpeg, .webp, .gif, .heic, .heif) appears anywhere in the conversation (typed by the user, injected as a `[Image: source: <path>]` line, or inside a tag) and you cannot see that image's content, run this skill on it before any other approach: no self-built OCR, no PIL, no tesseract. Also triggers on pasted-image placeholders such as `[Image #1]` and `[Unsupported Image]`. If you can actually see the image, do not use this skill. When unsure, run `modlens guard` before the first read of a session: a deny verdict means the active model has native vision and must read the image itself. Runs the modlens CLI to convert the image into structured JSON evidence: every word transcribed, layout regions, semantics, visual clues. Also use when the user asks how to install, configure, or switch modlens providers (Gemini API key, OpenAI-compatible endpoints, Claude API or Claude Code CLI).

3.3K
Stars
59/100
Trust
Kategorie: researchAudit

Fork of tess-two rewritten from scratch to support latest version of Tesseract OCR.

937
Stars
70/100
Trust
Kategorie: document-processingAudit