A command-line tool to crawl websites using puppeteer.
Supply asset profile
Deep research, source comparison, literature review, RAG, knowledge search, and reports.
Scenario
RAG and knowledge
I need my agent to build a RAG workflow over documents and retrieve reliable context.
Agent fit
Claude Code + Browser agents + CLI
Codex, Claude Code, Cursor, CLI, or custom agents.
Install
Ready
npx skills add patrickschur/pappet
Maintenance
stale
4y since push
Risk
Risky
Permission surface may require sandboxing
GitHub quality
105
46/100 quality · 65/100 trust
Coverage tags
Review notes
Permission surface may require sandboxing · Repository appears stale
Agent adoption scorecard
These scores combine public repository metadata, OpenAgentSkill review signals, maintenance freshness, and install readiness. They are a shortlist signal, not a replacement for human review.
Quality
Needs reviewInspect the repository carefully before adding it to an agent workflow.
Trust
Do not auto-installTrust Score v5 found insufficient evidence for agent installation. Treat this as discovery material, not an executable recommendation.
Audit
RiskyInstall readiness, security metadata, maintenance, and adoption risk.
Trust Score v5
Choose a stronger alternative or inspect the source manually before any install attempt.
Stars
105 GitHub stars
Repo activity
105 stars, 7 forks
Maintenance
4y since push
License
MIT
Install
npx skills add patrickschur/pappet
Install safety
standard package or runtime install path
Permission surface
shell or command execution, filesystem or document access
Agent outcomes
No agent outcome data yet
Docs
Usable metadata, review docs
Risk summary
Install readiness
Agent-readable metadata
Use this block or the embedded JSON to decide whether an agent should install this skill, choose an alternative, or ask for human review first.
Suited tasks
Suited agents
Install decision
Trust and risk
Outcome loop
Do not use when
Agent safety v2
This skill should not be selected by an agent without explicit human security review.
Do not auto-install. Inspect the source, dependencies, and permission surface first.
high
Skill metadata references terminal, CLI, shell, subprocess, or command execution workflows.
medium
Skill may drive a browser or interact with web pages.
medium
Skill likely fetches remote pages, APIs, repositories, or external services.
medium
Skill may read or write project files, documents, generated artifacts, or local workspace state.
Install targets
Copy the registry command or an agent-specific install prompt for Codex, Claude Code, and Cursor.
Use the registry command when your workflow supports the OpenAgentSkill installer.
$ npx skills add patrickschur/pappetAgent resolve plan
The Resolve API returns the selected skill, alternatives, safety policy, audit notes, install target, and copy-paste prompt an agent can follow without scraping this page.
Resolve JSON
/api/agent/resolve?task=Use%20Pappet%20for%20an%20agent%20workflow&agent=codex&max_risk=medium
Resolve text
/api/agent/resolve?task=Use%20Pappet%20for%20an%20agent%20workflow&agent=codex&max_risk=medium&format=text
Install handoff
/api/skills/patrickschur-pappet/install
Agent should check
Copy prompt
Task: Use Pappet in this workspace.
Resolve first: https://www.openagentskill.com/api/agent/resolve?task=Use%20Pappet%20for%20an%20agent%20workflow&agent=codex&max_risk=medium
Review install handoff: https://www.openagentskill.com/api/skills/patrickschur-pappet/install
Install command: npx skills add patrickschur/pappet
Before running it, summarize audit warnings, required permissions, and the fallback skill if install is risky.Agent handoff
Use the public install endpoint to fetch the command, safety checklist, target prompts, and canonical links for this skill.
Install handoff
/api/skills/patrickschur-pappet/install
LLM text format
/api/skills/patrickschur-pappet/install?format=text
Find alternatives
/api/skills/search?q=Pappet&limit=3
Agent prompt
Use Pappet for this task. Review https://www.openagentskill.com/api/skills/patrickschur-pappet/install, then install with: npx skills add patrickschur/pappetRegistry metadata
This page exposes the same decision, trust, audit, use-case, and install signals through the Registry API, so agents can rank this skill without scraping the UI.
Manifest
/api/registry/manifest/patrickschur-pappet
LLM text
/api/registry/manifest/patrickschur-pappet?format=text
Install alias
/api/registry/install/patrickschur-pappet
Recommend
/api/registry/recommend?task=Use%20Pappet%20in%20an%20agent%20workflow&limit=3
Agent fit
Web scraping
Use-case tags
Platforms
JavaScript, Puppeteer, Claude Code, Browser agents
Audit report
Review install readiness, maintenance, trust, quality, and metadata warnings before adding this skill to an agent workflow.
Agent decision cockpit
Do a manual repository review before adding this to an agent workflow.
Role in stack
Needs validation
Primary fit
Web scraping
Trust label
Needs manual review
Install path
Command ready
Use when
Evidence
Review first
Implementation path
Trust profile
Trust Score v5 found insufficient evidence for agent installation. Treat this as discovery material, not an executable recommendation.
GitHub adoption
INFO105 GitHub stars
Stars/forks activity
CHECK105 stars, 7 forks; issue activity unavailable in current metadata
Recent maintenance
FIX4y since push
License clarity
PASSMIT
Good signals
Review before install
Recommended action
Choose a stronger alternative or inspect the source manually before any install attempt.
Quality profile
Inspect the repository carefully before adding it to an agent workflow.
Workflow fit
Collect structured data
I need my agent to scrape websites and extract structured data from pages.
Search private knowledge
I need my agent to build a RAG workflow over documents and retrieve reliable context.
Operate web apps
I need my agent to control a browser, fill forms, and verify web app workflows.
Stack fit
Scrape, clean, and reuse web data
A practical stack for agents that crawl public pages, extract clean content, normalize data, and hand it to downstream research or RAG workflows.
Ingest, retrieve, and cite
A stack for document-heavy agents that ingest files, create searchable knowledge, retrieve relevant context, and answer with grounded sources.
Operate and verify web apps
A stack for agents that navigate products, fill forms, take screenshots, and verify real user flows across web applications.
Alternative shortlist
Similar skills in this category, ranked with the same readiness and quality signals.
Playwright is a framework for Web Testing and Automation. It allows testing Chromium, Firefox and WebKit with a single API.
Crawlee—A web scraping and browser automation library for Node.js to build reliable crawlers. In JavaScript and TypeScript. Extract data for AI, LLMs, RAG, or GPTs. Download HTML, PDF, JPG, PNG, and other files from websites. Works with Puppeteer, Playwright, Cheerio, JSDOM, and raw HTTP. Both headful and headless mode. With proxy rotation.
Python version of the Playwright testing and automation library.
Proxy server to bypass Cloudflare protection
A command-line tool to crawl websites using puppeteer.
Imported by the skill-only GitHub discovery pipeline because it matches agent skill, automation, domain workflow, RAG, document-processing, data, finance, security, or developer-tool signals. Protocol-server projects are excluded from automated imports.
Frameworks & Tools
Decision snapshot
install command or GitHub repo available
Audit snapshot
Install and adoption review
Agent-proven evidence
Outcome reports after resolve, review, install, and one narrow run.
No agent outcome data yet. The first agent run can report success, setup needs, risk blocks, failure, or not-relevant through /api/agent/outcome.
Install
Free and open source. Review the audit before production use.
Growth loop
Scenario-led draft for Pappet, ready for a manual X post.
Most web agents fail in the boring part: messy pages, missing context, repeatable extraction. Pappet gives agents a cleaner path to browse, extract, and monitor web pages. 105 stars https://www.openagentskill.com/skills/patrickschur-pappet?ref=x #AIAgents
Listing + install path for Pappet: https://www.openagentskill.com/skills/patrickschur-pappet?ref=x Install: npx skills add patrickschur/pappet
Listing source
This listing was indexed from public sources and is not marked official until a maintainer claim is approved.
Attribution links to the public repository or creator profile. Creators can claim the listing to update ownership signals.
Claim this skillOwner claim
This community indexed listing is attributed to patrickschur but is not marked official yet. Claim it to add a verified owner signal and make future launch, install, and audit updates easier to trust.
Creator backlink kit
Show the canonical listing, current trust and audit signals, and real Agent Proven evidence where developers evaluate the repository.
[](https://www.openagentskill.com/skills/patrickschur-pappet)
[](https://www.openagentskill.com/skills/patrickschur-pappet)
[](https://www.openagentskill.com/skills/patrickschur-pappet/audit)
[](https://www.openagentskill.com/skills/patrickschur-pappet)patrickschur
@patrickschur
Share whether this skill looks useful for your agent workflow. Aggregated feedback improves rankings over time.
Do not auto-install
Playwright
Playwright is a framework for Web Testing and Automation. It allows testing Chromium, Firefox and WebKit with a single API.
91.3K stars · 0 installsCrawlee
Crawlee—A web scraping and browser automation library for Node.js to build reliable crawlers. In JavaScript and TypeScript. Extract data for AI, LLMs, RAG, or GPTs. Download HTML, PDF, JPG, PNG, and other files from websites. Works with Puppeteer, Playwright, Cheerio, JSDOM, and raw HTTP. Both headful and headless mode. With proxy rotation.
24.0K stars · 0 installsPlaywright Python
Python version of the Playwright testing and automation library.
14.8K stars · 0 installsFlareSolverr
Proxy server to bypass Cloudflare protection
14.5K stars · 0 installs