Skill rankings

Best testing and qa skills for AI agents

Find skills that help agents generate tests, run browser checks, inspect failures, validate APIs, and keep product flows reliable.

Shown: 7 · Candidates: 62

Compare top 4

Saved directory data is shown because current data is unavailable. Dates and metrics may be out of date.

  1. 01

    Webapp Testing

    Use Playwright to interact with and test local web applications, capture screenshots, debug UI behavior, and inspect browser logs.

    @anthropicsBrowser Automation163,076
    Review before use
    Ranking signals

    Signals use a 0–100 scale. They explain the shortlist; the ordering follows this list’s method, not any one score.

    Popularity
    100/100
    Quality
    100/100
    Freshness
    0/100
    Agent evidence
    0/100
    Evidence confidence
    0/100
    Install readiness
    100/100
    Task fit
    29/100
  2. 02

    Playwright

    Reliable browser automation and testing engine for web agent tasks.

    @microsoftBrowser Automation76,000
    Review before use
    Ranking signals

    Signals use a 0–100 scale. They explain the shortlist; the ordering follows this list’s method, not any one score.

    Popularity
    98/100
    Quality
    100/100
    Freshness
    0/100
    Agent evidence
    0/100
    Evidence confidence
    0/100
    Install readiness
    90/100
    Task fit
    26/100
  3. 03

    Playwright Browser Skill

    Automate a real browser from the terminal for navigation, form filling, snapshots, screenshots, extraction, and UI-flow debugging.

    @openaiBrowser Automation24,000
    Review before use
    Ranking signals

    Signals use a 0–100 scale. They explain the shortlist; the ordering follows this list’s method, not any one score.

    Popularity
    88/100
    Quality
    100/100
    Freshness
    0/100
    Agent evidence
    0/100
    Evidence confidence
    0/100
    Install readiness
    100/100
    Task fit
    24/100
  4. 04

    Browser Use

    Browser automation layer for agents that need to interact with websites.

    @browser-useBrowser Automation75,000
    Review before use
    Ranking signals

    Signals use a 0–100 scale. They explain the shortlist; the ordering follows this list’s method, not any one score.

    Popularity
    98/100
    Quality
    100/100
    Freshness
    0/100
    Agent evidence
    0/100
    Evidence confidence
    0/100
    Install readiness
    77/100
    Task fit
    17/100
  5. 05

    Portfolio Health Check

    Audit concentration, factor exposure, correlation, liquidity, and stress-test risks in an existing portfolio.

    @GeeksfinoFinance271
    Review before use
    Ranking signals

    Signals use a 0–100 scale. They explain the shortlist; the ordering follows this list’s method, not any one score.

    Popularity
    49/100
    Quality
    100/100
    Freshness
    0/100
    Agent evidence
    0/100
    Evidence confidence
    0/100
    Install readiness
    90/100
    Task fit
    15/100
  6. 06

    Simons Quant

    Evaluate systematic strategy ideas through signal testing, statistical validation, decay, execution costs, and model-risk review.

    @xuboyuebobbFinance811
    Review before use
    Ranking signals

    Signals use a 0–100 scale. They explain the shortlist; the ordering follows this list’s method, not any one score.

    Popularity
    58/100
    Quality
    100/100
    Freshness
    0/100
    Agent evidence
    0/100
    Evidence confidence
    0/100
    Install readiness
    90/100
    Task fit
    13/100
  7. 07

    RNSkill Practical Video Planning

    Plan a practical long-form video with test prompts, recording structure, narration beats, screen-recording steps, and production handoff.

    @PluviobyteVideo Creation797
    Review before use
    Ranking signals

    Signals use a 0–100 scale. They explain the shortlist; the ordering follows this list’s method, not any one score.

    Popularity
    58/100
    Quality
    100/100
    Freshness
    0/100
    Agent evidence
    0/100
    Evidence confidence
    0/100
    Install readiness
    90/100
    Task fit
    13/100

A shortlist from up to 480 directory candidates, not the entire registry. Stars belong to repositories. Signals are not safety guarantees or runtime verification.

How this list works

Matches task and source metadata, then considers quality, popularity and freshness. Source descriptions are not execution evidence.

Signals use a 0–100 scale. They explain the shortlist; the ordering follows this list’s method, not any one score.