OpenAgentSkillRegistry
ResolveSkillsTasksPacksCompareAPIDocs
Star0Submit Skill
Star0
Skills/web-automation/Stormcrawler

Stormcrawler

STRONG · 78
Community indexed

A scalable, mature and versatile web crawler based on Apache Storm

Downloads0
Stars981
Version1.0.0
Quality87/100 · Excellent
Trust78/100 · Review then install
Audit89/100 · Safe to try

Supply asset profile

Coding and developer agents

Code review, repo analysis, testing, CI, GitHub, DevOps, and developer workflow skills.

Browse track

Scenario

Coding agents

I need a coding agent that can understand a repository, edit code, and review pull requests.

Agent fit

Claude Code + CLI + Codex

Codex, Claude Code, Cursor, CLI, or custom agents.

Install

Ready

npx skills add apache/stormcrawler

Maintenance

fresh

27d since push

Risk

Safe to try

Quality score needs review

GitHub quality

981

87/100 quality · 83/100 trust

Coverage tags

CodingCoding agentsweb-automationcrawlerdata-extraction

Review notes

Quality score needs review

Agent adoption scorecard

Trust, audit, and install readiness at a glance

These scores combine public repository metadata, OpenAgentSkill review signals, maintenance freshness, and install readiness. They are a shortlist signal, not a replacement for human review.

Quality

Excellent
87

High-confidence pick with strong adoption and healthy maintenance signals.

Trust

Review then install
78

Good shortlist signal, but the agent should review audit notes, install policy, and outcome evidence before running it.

Audit

Safe to try
89

Install readiness, security metadata, maintenance, and adoption risk.

Trust Score v5

Human review before install

Use as the primary candidate after human or sandbox review.

JavaCrawlerCodexClaude CodeCursor

Stars

981 GitHub stars

Repo activity

981 stars, 282 forks

Maintenance

27d since push

License

Apache-2.0

Install

npx skills add apache/stormcrawler

Install safety

standard package or runtime install path

Permission surface

no high-risk permission surface in public metadata

Agent outcomes

No agent outcome data yet

Docs

Usable metadata, review docs

Risk summary

Low metadata risk

  • Quality score needs review

Install readiness

Install path available

  • Install path is available
  • Repository evidence is available
  • License is declared
  • No Agent Proven outcome evidence yet

Agent-readable metadata

Machine-readable decision data for this skill.

Use this block or the embedded JSON to decide whether an agent should install this skill, choose an alternative, or ask for human review first.

Open JSON

Suited tasks

  • Web scraping workflows
  • Claude Code teams
  • teams that value GitHub adoption signals
  • Crawl target URLs

Suited agents

JavaCrawlerCodexClaude CodeCursorOpenAgentSkill CLICLI

Install decision

Command
npx skills add apache/stormcrawler
Policy
review
Human review
yes

Trust and risk

Trust
78/100
Audit
89/100
Risk level
Safe to try

Outcome loop

Endpoint
/api/agent/outcome
Event ID
resolve
Outcomes
5

Install command

npx skills add apache/stormcrawler
Public auditEval reportResolve APIInstall handoff

Do not use when

  • teams that need a vendor-supported SLA
  • high-compliance environments without internal security review
  • No major risk signals from current metadata
  • Quality score needs review
  • Production credentials, payments, or irreversible account changes without explicit human review

Alternative

Crawl4AI

73.1K stars

npx skills add unclecode/crawl4ai

Alternative

Scrapling

70.0K stars

npx skills add D4Vinci/Scrapling

Alternative

Scrapy

62.5K stars

npx skills add scrapy/scrapy

Alternative

EasySpider

44.1K stars

npx skills add NaiboWang/EasySpider

Agent safety v2

77/100 · Review before install

Reviewedreview

Good audit and safety signals with no high-risk permission hints in public metadata.

Review the audit page, then allow agent install in a sandboxed workflow.

Resolve via API

medium

Network access

Skill likely fetches remote pages, APIs, repositories, or external services.

  • Quality score needs review

Install targets

Install this skill in your agent workflow

Copy the registry command or an agent-specific install prompt for Codex, Claude Code, and Cursor.

skill install

OpenAgentSkill CLI

Use the registry command when your workflow supports the OpenAgentSkill installer.

$ npx skills add apache/stormcrawler

Agent resolve plan

Let an agent verify fit before installing.

The Resolve API returns the selected skill, alternatives, safety policy, audit notes, install target, and copy-paste prompt an agent can follow without scraping this page.

Open text plan

Resolve JSON

/api/agent/resolve?task=Use%20Stormcrawler%20for%20an%20agent%20workflow&agent=codex&max_risk=medium

Resolve text

/api/agent/resolve?task=Use%20Stormcrawler%20for%20an%20agent%20workflow&agent=codex&max_risk=medium&format=text

Install handoff

/api/skills/apache-stormcrawler/install

Agent should check

  • Task fit and alternatives from Resolve API.
  • Audit score, trust score, and safety policy warnings.
  • Install target compatibility for Codex, Claude Code, Cursor, or CLI.

Copy prompt

Task: Use Stormcrawler in this workspace.
Resolve first: https://www.openagentskill.com/api/agent/resolve?task=Use%20Stormcrawler%20for%20an%20agent%20workflow&agent=codex&max_risk=medium
Review install handoff: https://www.openagentskill.com/api/skills/apache-stormcrawler/install
Install command: npx skills add apache/stormcrawler
Before running it, summarize audit warnings, required permissions, and the fallback skill if install is risky.

Agent handoff

Give an agent the install path, not another directory page.

Use the public install endpoint to fetch the command, safety checklist, target prompts, and canonical links for this skill.

Open install API

Install handoff

/api/skills/apache-stormcrawler/install

LLM text format

/api/skills/apache-stormcrawler/install?format=text

Find alternatives

/api/skills/search?q=Stormcrawler&limit=3

Agent prompt

Use Stormcrawler for this task. Review https://www.openagentskill.com/api/skills/apache-stormcrawler/install, then install with: npx skills add apache/stormcrawler

Registry metadata

Agent-readable profile for automatic skill selection.

This page exposes the same decision, trust, audit, use-case, and install signals through the Registry API, so agents can rank this skill without scraping the UI.

Open manifest

Manifest

/api/registry/manifest/apache-stormcrawler

LLM text

/api/registry/manifest/apache-stormcrawler?format=text

Install alias

/api/registry/install/apache-stormcrawler

Recommend

/api/registry/recommend?task=Use%20Stormcrawler%20in%20an%20agent%20workflow&limit=3

Agent fit

100/100

Web scraping

Use-case tags

Web scrapingCoding agentsSports analytics

Platforms

Java, Crawler, Claude Code

Audit report

Safe to try · 89/100

Review install readiness, maintenance, trust, quality, and metadata warnings before adding this skill to an agent workflow.

View audit reportView eval report

Agent decision cockpit

Primary pick for Web scraping

Use this as a leading candidate, then validate the README and install path in your own agent stack.

100
Readiness
Adopt
Stage

Role in stack

Primary pick

Primary fit

Web scraping

Trust label

Production-ready

Install path

Command ready

Use when

  • Web scraping workflows
  • Claude Code teams
  • teams that value GitHub adoption signals

Evidence

  • 981 GitHub stars
  • recent repository activity
  • install command or GitHub repo available
  • 87/100 quality profile
  • 8 OpenAgentSkill engagement events

Review first

  • No major risk signals from current metadata

Implementation path

  1. 1Install it in a sandbox agent and run one Web scraping task end to end.
  2. 2Compare output quality, latency, and failure behavior against at least one alternative.
  3. 3Promote it into production only after reviewing repository permissions, license, and maintenance signals.

Trust profile

Review then install

Good shortlist signal, but the agent should review audit notes, install policy, and outcome evidence before running it.

78
Trust score

GitHub adoption

INFO

981 GitHub stars

Stars/forks activity

INFO

981 stars, 282 forks; issue activity unavailable in current metadata

Recent maintenance

PASS

27d since push

License clarity

PASS

Apache-2.0

Good signals

  • AI review approved
  • Install path is available
  • Repository evidence is available
  • Recently maintained repository
  • Meaningful GitHub adoption signal
  • Install command has no obvious high-risk pattern
  • Outcome loop is ready but needs first real agent run

Review before install

  • Quality score needs review
  • No real agent outcome reports yet
  • Human review required before unattended installation

Recommended action

Use as the primary candidate after human or sandbox review.

Quality profile

Excellent candidate for agent workflows

High-confidence pick with strong adoption and healthy maintenance signals.

87
GitHub stars
981
Freshness
27d ago
Install ready
Yes
License
Apache-2.0

Workflow fit

Use this skill in these scenarios

Collect structured data

Web scraping

I need my agent to scrape websites and extract structured data from pages.

Build and ship code

Coding agents

I need a coding agent that can understand a repository, edit code, and review pull requests.

Analyze matches

Sports analytics

I need my agent to analyze football matches, World Cup data, xG, players, teams, and predictions.

Stack fit

Add it to a complete workflow

Scrape, clean, and reuse web data

Web data pipeline

A practical stack for agents that crawl public pages, extract clean content, normalize data, and hand it to downstream research or RAG workflows.

Inspect, patch, and verify code

Coding review agent

A stack for software agents that inspect repositories, review pull requests, generate tests, and turn findings into shippable patches.

Turn skills into distribution

Content growth agent

A stack for turning newly indexed skills into SEO briefs, social drafts, comparison pages, and reusable publishing workflows.

Alternative shortlist

Compare before you install

Similar skills in this category, ranked with the same readiness and quality signals.

Compare all

Crawl4AI

Open-source LLM-friendly web crawler and scraper

Adopt
Ready100
Quality100
Stars73.1K

Scrapling

Adaptive web scraping for agent data collection

Adopt
Ready100
Quality100
Stars70.0K

Scrapy

Scrapy, a fast high-level web crawling & scraping framework for Python.

Adopt
Ready100
Quality100
Stars62.5K

EasySpider

A visual no-code/code-free web crawler/spider易采集:一个可视化浏览器自动化测试/数据采集/网页爬虫软件,可以无代码图形化的设计和执行爬虫任务。别名:ServiceWrapper面向Web应用的智能化服务封装系统。

Adopt
Ready100
Quality100
Stars44.1K

Overview

A scalable, mature and versatile web crawler based on Apache Storm

Imported by the skill-only GitHub discovery pipeline because it matches agent skill, automation, RAG, or developer-tool signals. Protocol-server projects are excluded from automated imports.

Platform Compatibility

javaFULL
crawlerFULL

Technical Details

Version
1.0.0
License
Apache-2.0
Last Updated
6/25/2026
Published
5/23/2026

Frameworks & Tools

JavaCrawler

Decision snapshot

Primary pick

100
Ready
Adopt
Stage

981 GitHub stars

Audit snapshot

Install review

Install and adoption review

89
Safe to try
Security
89/100
Maintenance
100/100
Install
92/100
Open full auditOpen eval report

Agent-proven evidence

Agent Proven evidence

Outcome reports after resolve, review, install, and one narrow run.

0
Proven
Needs first agent runAuto-install: review firstLast: Unknown
Success rate
—
Recent failure
—
Outcomes
0
Output quality
—
Failed
0
Not relevant
0
Installs
0
Risk blocked
0
Setup needed
0
Production
0

No agent outcome data yet. The first agent run can report success, setup needs, risk blocks, failure, or not-relevant through /api/agent/outcome.

Agent-Proven rankingOutcome contract

Install

Add to agent workflow

Free and open source. Review the audit before production use.

Compare AlternativesAuto-resolve PlanView on GitHubDocumentation

Growth loop

Share kit

X

Scenario-led draft for Stormcrawler, ready for a manual X post.

Curator note
Most web agents fail in the boring part: messy pages, missing context, repeatable extraction.

Stormcrawler gives agents a cleaner path to browse, extract, and monitor web pages.

981 stars

https://www.openagentskill.com/skills/apache-stormcrawler?ref=x
#AIAgents
Open X draft
Optional reply with install command
Listing + install path for Stormcrawler:
https://www.openagentskill.com/skills/apache-stormcrawler?ref=x

Install: npx skills add apache/stormcrawler
Open reply draft

Listing source

Community indexed

Claimable

This listing was indexed from public sources and is not marked official until a maintainer claim is approved.

Creator
apache
Source
apache/stormcrawler
Indexed by
OpenAgentSkill community index

Attribution links to the public repository or creator profile. Creators can claim the listing to update ownership signals.

Claim this skill

Owner claim

Claim this skill listing

This community indexed listing is attributed to apache but is not marked official yet. Claim it to add a verified owner signal and make future launch, install, and audit updates easier to trust.

Creator backlink kit

Add the evidence badges to your README

Show the canonical listing, current trust and audit signals, and real Agent Proven evidence where developers evaluate the repository.

[![Listed on OpenAgentSkill](https://www.openagentskill.com/api/badge/apache-stormcrawler?metric=listed&label=Listed)](https://www.openagentskill.com/skills/apache-stormcrawler)
[![OpenAgentSkill Trust](https://www.openagentskill.com/api/badge/apache-stormcrawler?metric=trust&label=Trust)](https://www.openagentskill.com/skills/apache-stormcrawler)
[![OpenAgentSkill Audit](https://www.openagentskill.com/api/badge/apache-stormcrawler?metric=audit&label=Audit)](https://www.openagentskill.com/skills/apache-stormcrawler/audit)
[![Agent Proven](https://www.openagentskill.com/api/badge/apache-stormcrawler?metric=proven&label=Agent%20Proven)](https://www.openagentskill.com/skills/apache-stormcrawler)
Preview badge Open audit Creator Kit

Author

A

apache

@apache

Tags

crawlerdata-extractionapache-stormdistributedjavastormcrawlerweb-crawlergithub

Platform Fit

Claude Code

Health Signals

GitHub stars
981
Quality score
55/100
Last GitHub push
Jun 24, 2026
Framework hints
2
OpenAgentSkill views
3
Install copies
0
Outbound clicks
0

Community Signal

Share whether this skill looks useful for your agent workflow. Aggregated feedback improves rankings over time.

Trust & Safety

Review then install

78
  • GitHub adoption981 GitHub starsINFO
  • Stars/forks activity981 stars, 282 forks; issue activity unavailable in current metadataINFO
  • Recent maintenance27d since pushPASS
  • License clarityApache-2.0PASS
  • README/SKILL.md completenessPublic metadata needs stronger README/SKILL.md contextINFO
  • Dependency/runtime riskno major dependency risk hints in public metadataPASS

Related Skills

Crawl4AI

Open-source LLM-friendly web crawler and scraper

73.1K stars · 31.0K installs

Scrapling

Adaptive web scraping for agent data collection

70.0K stars · 0 installs

Scrapy

Scrapy, a fast high-level web crawling & scraping framework for Python.

62.5K stars · 0 installs

EasySpider

A visual no-code/code-free web crawler/spider易采集:一个可视化浏览器自动化测试/数据采集/网页爬虫软件,可以无代码图形化的设计和执行爬虫任务。别名:ServiceWrapper面向Web应用的智能化服务封装系统。

44.1K stars · 0 installs
OpenAgentSkill

The skill layer for AI agents: discover, compare, audit, and install reusable capabilities across Codex, Claude Code, Cursor, and agent runtimes.

GitHubX

Explore

SkillsAgent SkillsWhat Is an Agent Skill?AI Agent SkillsTasksSkill PacksBest SkillsTrendingStacksUse CasesAgentsAgent Entry

Trust

CompareSafety GateSkills RegistryRankingsOutcomesAuditsOfficialWeekly ReportsMonthly IndexState of Agent Skillsvs skills.shAgentSkills.io Alternative

Build

DocsAbout OpenAgentSkillAPIllms.txtOpenAPICLICreator KitX Growth KitSubmitBlogGuidesActivity
OpenAgentSkill Registry
PrivacyBuilt for agent-native discovery