Skill comparison
Use this as a shortlist, then open the skill detail page before adopting.
Decision summary
Strongest overall
promptfoo-evals
Use this as a leading candidate, then validate the README and install path in your own agent stack.
Fastest prototype
promptfoo-evals
Best first install candidate based on install readiness and adoption.
Freshest repo
promptfoo-evals
Most recent maintenance signal among this shortlist.
| Signal | promptfoo-evals Write, refine, run, and QA non-redteam promptfoo eval suites after the target or provider already works: prompts, vars, test cases, assertions, model-graded rubrics, transforms, datasets, output exports, filters, and CI gates. Use for regression tests and eval-suite authoring. Do not use for connecting a new target/provider, mapping HTTP requests or auth, smoke-testing an endpoint, or redteam plugin/strategy setup; use `promptfoo-provider-setup` for connection work instead. |
|---|---|
| Quality | 91/100 Excellent |
| Decision verdict | 100/100 Production-ready Use this as a leading candidate, then validate the README and install path in your own agent stack. |
| Adoption | 25K stars Verified outcomes are shown on each skill page |
| Freshness | Sep 2, 2026 |
| Use-case fit | |
| Workflow fit | |
| Platform hints | Claude Code, OpenAI Agents |
| Warnings | No OpenAgentSkill engagement data yet |
| Best for |
Skill comparison
Use this as a shortlist, then open the skill detail page before adopting.
Decision summary
Strongest overall
promptfoo-evals
Use this as a leading candidate, then validate the README and install path in your own agent stack.
Fastest prototype
promptfoo-evals
Best first install candidate based on install readiness and adoption.
Freshest repo
promptfoo-evals
Most recent maintenance signal among this shortlist.
| Signal | promptfoo-evals Write, refine, run, and QA non-redteam promptfoo eval suites after the target or provider already works: prompts, vars, test cases, assertions, model-graded rubrics, transforms, datasets, output exports, filters, and CI gates. Use for regression tests and eval-suite authoring. Do not use for connecting a new target/provider, mapping HTTP requests or auth, smoke-testing an endpoint, or redteam plugin/strategy setup; use `promptfoo-provider-setup` for connection work instead. |
|---|---|
| Quality | 91/100 Excellent |
| Decision verdict | 100/100 Production-ready Use this as a leading candidate, then validate the README and install path in your own agent stack. |
| Adoption | 25K stars Verified outcomes are shown on each skill page |
| Freshness | Sep 2, 2026 |
| Use-case fit | |
| Workflow fit | |
| Platform hints | Claude Code, OpenAI Agents |
| Warnings | No OpenAgentSkill engagement data yet |
| Best for |
| Research agents workflows · Claude Code teams · teams that value GitHub adoption signals |
| Not ideal for | teams that need a vendor-supported SLA · high-compliance environments without internal security review |
| OpenAgentSkill engagement | 0 views 0 install copies |
| Install | $ npx skills add promptfoo/promptfoo --skill promptfoo-evals |
| Research agents workflows · Claude Code teams · teams that value GitHub adoption signals |
| Not ideal for | teams that need a vendor-supported SLA · high-compliance environments without internal security review |
| OpenAgentSkill engagement | 0 views 0 install copies |
| Install | $ npx skills add promptfoo/promptfoo --skill promptfoo-evals |