Direktori skill

Temukan skill yang dapat digunakan kembali untuk AI agents.

Cari skill GitHub nyata berdasarkan tugas lalu periksa stars, trust, audit, kategori, dan jalur pemasangan sebelum digunakan.

Setiap rekomendasi tetap terhubung dengan repositori, audit, dan jalur pemasangannya.

Hasil pencarian: tesseract

Direktori bahasa Inggris

ICCV 2025 | TesserAct: Learning 4D Embodied World Models

400
Stars
62/100
Kepercayaan
Kategori: robotics-iotAudit

Motion Planning Environment

367
Stars
65/100
Kepercayaan
Kategori: robotics-iotAudit

Tesseract Open Source OCR Engine (main repository)

75K
Stars
81/100
Kepercayaan
Kategori: document-processingAudit

Pure Javascript OCR for more than 100 Languages ๐Ÿ“–๐ŸŽ‰๐Ÿ–ฅ

38K
Stars
81/100
Kepercayaan
Kategori: document-processingAudit

Enhances Tesseract OCR output using LLMs (local or API) for error correction, smart chunking, and markdown formatting of scanned PDFs

2.9K
Stars
73/100
Kepercayaan
Kategori: document-processingAudit

A Python wrapper for the tesseract-ocr API

2.2K
Stars
75/100
Kepercayaan
Kategori: document-processingAudit

Go package for OCR (Optical Character Recognition), by using Tesseract C++ library

3.1K
Stars
79/100
Kepercayaan
Kategori: document-processingAudit

A wrapper to work with Tesseract OCR inside PHP.

3.0K
Stars
75/100
Kepercayaan
Kategori: document-processingAudit

Trained models with fast variant of the "best" LSTM models + legacy models

7.6K
Stars
71/100
Kepercayaan
Kategori: document-processingAudit

A Gtk/Qt front-end to tesseract-ocr.

2.0K
Stars
73/100
Kepercayaan
Kategori: document-processingAudit

Plug-in vision for text-only models. Hard rule: when a file path or URL with an image extension (.png, .jpg, .jpeg, .webp, .gif, .heic, .heif) appears anywhere in the conversation (typed by the user, injected as a `[Image: source: <path>]` line, or inside a tag) and you cannot see that image's content, run this skill on it before any other approach: no self-built OCR, no PIL, no tesseract. Also triggers on pasted-image placeholders such as `[Image #1]` and `[Unsupported Image]`. If you can actually see the image, do not use this skill. When unsure, run `modlens guard` before the first read of a session: a deny verdict means the active model has native vision and must read the image itself. Runs the modlens CLI to convert the image into structured JSON evidence: every word transcribed, layout regions, semantics, visual clues. Also use when the user asks how to install, configure, or switch modlens providers (Gemini API key, OpenAI-compatible endpoints, Claude API or Claude Code CLI).

3.3K
Stars
59/100
Kepercayaan
Kategori: researchAudit

Fork of tess-two rewritten from scratch to support latest version of Tesseract OCR.

937
Stars
70/100
Kepercayaan
Kategori: document-processingAudit