Skill rankings
Best testing and qa skills for AI agents
Find skills that help agents generate tests, run browser checks, inspect failures, validate APIs, and keep product flows reliable.
Shown: 7 · Candidates: 62
Compare top 4Saved directory data is shown because current data is unavailable. Dates and metrics may be out of date.
- 01
Webapp Testing
Use Playwright to interact with and test local web applications, capture screenshots, debug UI behavior, and inspect browser logs.
@anthropicsBrowser Automation163,076Review before useRanking signals
Signals use a 0–100 scale. They explain the shortlist; the ordering follows this list’s method, not any one score.
- Popularity
- 100/100
- Quality
- 100/100
- Freshness
- 0/100
- Agent evidence
- 0/100
- Evidence confidence
- 0/100
- Install readiness
- 100/100
- Task fit
- 29/100
- 02
Playwright
Reliable browser automation and testing engine for web agent tasks.
@microsoftBrowser Automation76,000Review before useRanking signals
Signals use a 0–100 scale. They explain the shortlist; the ordering follows this list’s method, not any one score.
- Popularity
- 98/100
- Quality
- 100/100
- Freshness
- 0/100
- Agent evidence
- 0/100
- Evidence confidence
- 0/100
- Install readiness
- 90/100
- Task fit
- 26/100
- 03
Playwright Browser Skill
Automate a real browser from the terminal for navigation, form filling, snapshots, screenshots, extraction, and UI-flow debugging.
@openaiBrowser Automation24,000Review before useRanking signals
Signals use a 0–100 scale. They explain the shortlist; the ordering follows this list’s method, not any one score.
- Popularity
- 88/100
- Quality
- 100/100
- Freshness
- 0/100
- Agent evidence
- 0/100
- Evidence confidence
- 0/100
- Install readiness
- 100/100
- Task fit
- 24/100
- 04
Browser Use
Browser automation layer for agents that need to interact with websites.
@browser-useBrowser Automation75,000Review before useRanking signals
Signals use a 0–100 scale. They explain the shortlist; the ordering follows this list’s method, not any one score.
- Popularity
- 98/100
- Quality
- 100/100
- Freshness
- 0/100
- Agent evidence
- 0/100
- Evidence confidence
- 0/100
- Install readiness
- 77/100
- Task fit
- 17/100
- 05
Portfolio Health Check
Audit concentration, factor exposure, correlation, liquidity, and stress-test risks in an existing portfolio.
@GeeksfinoFinance271Review before useRanking signals
Signals use a 0–100 scale. They explain the shortlist; the ordering follows this list’s method, not any one score.
- Popularity
- 49/100
- Quality
- 100/100
- Freshness
- 0/100
- Agent evidence
- 0/100
- Evidence confidence
- 0/100
- Install readiness
- 90/100
- Task fit
- 15/100
- 06
Simons Quant
Evaluate systematic strategy ideas through signal testing, statistical validation, decay, execution costs, and model-risk review.
@xuboyuebobbFinance811Review before useRanking signals
Signals use a 0–100 scale. They explain the shortlist; the ordering follows this list’s method, not any one score.
- Popularity
- 58/100
- Quality
- 100/100
- Freshness
- 0/100
- Agent evidence
- 0/100
- Evidence confidence
- 0/100
- Install readiness
- 90/100
- Task fit
- 13/100
- 07
RNSkill Practical Video Planning
Plan a practical long-form video with test prompts, recording structure, narration beats, screen-recording steps, and production handoff.
@PluviobyteVideo Creation797Review before useRanking signals
Signals use a 0–100 scale. They explain the shortlist; the ordering follows this list’s method, not any one score.
- Popularity
- 58/100
- Quality
- 100/100
- Freshness
- 0/100
- Agent evidence
- 0/100
- Evidence confidence
- 0/100
- Install readiness
- 90/100
- Task fit
- 13/100
A shortlist from up to 480 directory candidates, not the entire registry. Stars belong to repositories. Signals are not safety guarantees or runtime verification.
How this list works
Matches task and source metadata, then considers quality, popularity and freshness. Source descriptions are not execution evidence.
Signals use a 0–100 scale. They explain the shortlist; the ordering follows this list’s method, not any one score.













