n8n-binary-and-data
czlonkowski
Handle files and binary data in n8n correctly. Use when working with files, images, PDFs, attachments, uploads or downloads, base64, vision/multimodal input, or when an…
OPENAGENTSKILL / DIRECTORY
Find a skill for your next task. Explore tools for Codex, Claude Code, Cursor and more.
Find a skill for your next task. Preview examples where available.
145–160 / 168
Results: 168
czlonkowski
Handle files and binary data in n8n correctly. Use when working with files, images, PDFs, attachments, uploads or downloads, base64, vision/multimodal input, or when an…
danfickle
An HTML to PDF library for the JVM. Based on Flying Saucer and Apache PDF-BOX 2. With SVG image support. Now also with accessible PDF support (WCAG, Section 508, PDF/UA)!
bgshih
Convolutional Recurrent Neural Network (CRNN) for image-based sequence recognition.
UniversalViewer
A community-developed open source project on a mission to help you share your 📚📜📰📽️📻🗿 with the 🌎
realskyrin
⌘⌘ - A lightweight, native macOS screenshot tool that lives in your menu bar. Double-tap ⌘ Command to capture any region of your screen — instantly copied to clipboard,…
thanhkeke97
🎮 Real-time Game Translation Tool | OCR + AI Translation | Windows Gaming | Open Source
Belval
A python module that wraps the pdftoppm utility to convert PDF to PIL Image object
clawsoftware
Open Source Virtual (Network) Printer for Windows that allows you to create PDFs, OCR text, and print images, with advanced features usually available only in enterprise…
yakovmeister
A utility for converting pdf to image and base64 format.
Sierkinhane
(CRNN) Chinese Characters Recognition.
run-llama
ParseBench - A Document Parsing Benchmark for AI Agents
ianzhao
Python tool for grabbing text via screenshot
rtr46
meikipop - universal japanese ocr popup dictionary for windows, linux and macos
vorojar
Open-source batch OCR workbench — a free, local alternative to ABBYY FineReader. Powered by Ollama + GLM-OCR + PP-DocLayoutV3, ~0.5s/page on RTX 4090. Three-panel editor…
heyderekj
RQLuo
MixTeX multimodal LaTeX, ZhEn, and, Table OCR. It performs efficient CPU-based inference in a local offline on Windows.
Design direction, Figma implementation, UI review, React performance, browser QA, and safe preview deployments.
Code review, repo inspection, testing, planning, shipping, and engineering workflows for coding agents.
Research, recent-events briefings, PDF parsing, markdown conversion, RAG ingestion, and knowledge workflows.
Crawling, scraping, extraction, browser automation, structured data capture, and website-to-markdown workflows.
Image, video, creative production, UI design, multimodal generation, and visual workflow skills.
Data analysis, analytics, ETL, notebooks, databases, tables, charts, and reporting workflows.
SEO, content research, growth workflows, social listening, campaign analysis, and publishing helpers.
Teaching, tutoring, course creation, lesson planning, learning support, and explanation skills.