Directorio de skills

Descubre skills reutilizables para AI agents.

Busca skills reales de GitHub por tarea y revisa stars, confianza, auditoría, categoría y ruta de instalación antes de utilizarlos.

Cada recomendación conserva un vínculo claro con su repositorio, auditoría y ruta de instalación.

Resultados de búsqueda: tesseract

Directorio en inglés

Tesseract Open Source OCR Engine (main repository)

75K
Stars
81/100
Confianza
Categoría: document-processingAuditoría

ICCV 2025 | TesserAct: Learning 4D Embodied World Models

400
Stars
62/100
Confianza
Categoría: robotics-iotAuditoría

Motion Planning Environment

367
Stars
65/100
Confianza
Categoría: robotics-iotAuditoría

Pure Javascript OCR for more than 100 Languages 📖🎉🖥

38K
Stars
81/100
Confianza
Categoría: document-processingAuditoría

Enhances Tesseract OCR output using LLMs (local or API) for error correction, smart chunking, and markdown formatting of scanned PDFs

2.9K
Stars
73/100
Confianza
Categoría: document-processingAuditoría

A Python wrapper for the tesseract-ocr API

2.2K
Stars
75/100
Confianza
Categoría: document-processingAuditoría

Go package for OCR (Optical Character Recognition), by using Tesseract C++ library

3.1K
Stars
79/100
Confianza
Categoría: document-processingAuditoría

Trained models with fast variant of the "best" LSTM models + legacy models

7.6K
Stars
71/100
Confianza
Categoría: document-processingAuditoría

A Gtk/Qt front-end to tesseract-ocr.

2.0K
Stars
73/100
Confianza
Categoría: document-processingAuditoría

Plug-in vision for text-only models. Hard rule: when a file path or URL with an image extension (.png, .jpg, .jpeg, .webp, .gif, .heic, .heif) appears anywhere in the conversation (typed by the user, injected as a `[Image: source: <path>]` line, or inside a tag) and you cannot see that image's content, run this skill on it before any other approach: no self-built OCR, no PIL, no tesseract. Also triggers on pasted-image placeholders such as `[Image #1]` and `[Unsupported Image]`. If you can actually see the image, do not use this skill. When unsure, run `modlens guard` before the first read of a session: a deny verdict means the active model has native vision and must read the image itself. Runs the modlens CLI to convert the image into structured JSON evidence: every word transcribed, layout regions, semantics, visual clues. Also use when the user asks how to install, configure, or switch modlens providers (Gemini API key, OpenAI-compatible endpoints, Claude API or Claude Code CLI).

3.3K
Stars
59/100
Confianza
Categoría: researchAuditoría

Fork of tess-two rewritten from scratch to support latest version of Tesseract OCR.

937
Stars
70/100
Confianza
Categoría: document-processingAuditoría