Directorio de skills

Descubre skills reutilizables para AI agents.

Busca skills reales de GitHub por tarea y revisa stars, confianza, auditoría, categoría y ruta de instalación antes de utilizarlos.

Cada recomendación conserva un vínculo claro con su repositorio, auditoría y ruta de instalación.

Resultados de búsqueda: extraction

Directorio en inglés

Automate a real browser from the terminal for navigation, form filling, snapshots, screenshots, extraction, and UI-flow debugging.

25K
Stars
73/100
Confianza
Categoría: browser-automationAuditoría

🔥 The open-source no-code platform for web scraping, crawling, search and AI data extraction • Turn websites into structured APIs in minutes 🔥

16K
Stars
85/100
Confianza
Categoría: web-automationAuditoría

newspaper3k is a news, full-text, and article metadata extraction in Python 3. Advanced docs:

15K
Stars
87/100
Confianza
Categoría: web-automationAuditoría

Python & Command-line tool to gather text and metadata on the Web: Crawling, scraping, extraction, output as CSV, JSON, HTML, MD, TXT, XML

6.4K
Stars
85/100
Confianza
Categoría: web-automationAuditoría

PyMuPDF is a high performance Python library for data extraction, analysis, conversion & manipulation of PDF (and other) documents.

10K
Stars
86/100
Confianza
Categoría: document-processingAuditoría

Structured data extraction and instruction calling with ML, LLM and Vision LLM

5.2K
Stars
81/100
Confianza
Categoría: dataAuditoría

A tool to convert a Wallpaper's color scheme / palette, OCR with VLM's Traditional & Hybrid, Image Compression ,color palette extraction, image upsacling with Adversarial Networks and more image processing features.

2.3K
Stars
79/100
Confianza
Categoría: document-processingAuditoría

🕵️‍♂️ (2-in-1) Email & Username OSINT suite for deep data extraction. Analyzes 240+ scan vectors (100+ email / 140+ username) for security research, investigations, and digital footprinting.

2.2K
Stars
84/100
Confianza
Categoría: securityAuditoría

A full-lifecycle Claude Code skill for creating, compiling, reviewing, and polishing academic Beamer LaTeX presentations with quality scoring and pedagogical audits.

324
Stars
77/100
Confianza
Categoría: presentationAuditoría

Synthetic data curation for post-training and structured data extraction

1.7K
Stars
77/100
Confianza
Categoría: ml-automationAuditoría

Fast Rust library for PDF inspection, classification, and text extraction. Intelligently detects scanned vs text-based PDFs to enable smart routing decisions.

1.5K
Stars
75/100
Confianza
Categoría: document-processingAuditoría

QuantMind is an intelligent knowledge extraction and retrieval framework for quantitative finance.

1.4K
Stars
79/100
Confianza
Categoría: financeAuditoría