Deepseek Ocr.Rs
TimmyOVO
Rust multi‑backend OCR/VLM engine (DeepSeek‑OCR-1/2, PaddleOCR‑VL, DotsOCR) with DSQ quantization and an OpenAI‑compatible server & CLI – run locally without Python.
OPENAGENTSKILL / DIRECTORY
Find a skill for your next task. Explore tools for Codex, Claude Code, Cursor and more.
Find a skill for your next task. Preview examples where available.
225–240 / 481
Results: 481
TimmyOVO
Rust multi‑backend OCR/VLM engine (DeepSeek‑OCR-1/2, PaddleOCR‑VL, DotsOCR) with DSQ quantization and an OpenAI‑compatible server & CLI – run locally without Python.
adithya-s-k
Ingest, parse, and optimize any data format ➡️ from documents to multimedia ➡️ for enhanced compatibility with GenAI frameworks
NanoNets
An on-premises, OCR-free unstructured data extraction, markdown conversion and benchmarking toolkit. (https://idp-leaderboard.org/)
lukas-blecher
pix2tex: Using a ViT to convert images of equations into LaTeX code.
salomonelli
:necktie: :briefcase: Build fast :rocket: and easy multiple beautiful resumes and create your best CV ever! Made with Vue and LESS.
LetTTGACO
Markdown 批量导出工具、开放式跨平台博客解决方案,随意组合写作平台(语雀/Notion/FlowUs/飞书/我来Wolai)和博客平台(Hexo/Vitepress/Halo/Confluence/WordPress等)
Lulzx
Minimal PDF creation library. <400 LOC, zero dependencies, makes real PDFs.
AlibabaResearch
A collection of original, innovative ideas and algorithms towards Advanced Literate Machinery. This project is maintained by the OCR Team in the Language Technology Lab,…
camelot-dev
A web interface to extract tabular data from PDFs
sile-typesetter
The SILE Typesetter — Simon’s Improved Layout Engine
jesselau76
Enjoy reading with your favorite style.
pandao
The open source embeddable online markdown editor (component).
baskerville
axa-group
Transforms PDF, Documents and Images into Enriched Structured Data
kha-white
Read Japanese manga inside browser with selectable text.
nazdridoy
A CLI text-to-speech tool using the Kokoro model, supporting multiple languages, voices (with blending), and various input formats including EPUB books and PDF documents.
Design direction, Figma implementation, UI review, React performance, browser QA, and safe preview deployments.
Code review, repo inspection, testing, planning, shipping, and engineering workflows for coding agents.
Research, recent-events briefings, PDF parsing, markdown conversion, RAG ingestion, and knowledge workflows.
Crawling, scraping, extraction, browser automation, structured data capture, and website-to-markdown workflows.
Image, video, creative production, UI design, multimodal generation, and visual workflow skills.
Data analysis, analytics, ETL, notebooks, databases, tables, charts, and reporting workflows.
SEO, content research, growth workflows, social listening, campaign analysis, and publishing helpers.
Teaching, tutoring, course creation, lesson planning, learning support, and explanation skills.