Skill comparison
Use this as a shortlist, then open the skill detail page before adopting.
Decision summary
Strongest overall
skill-upper
Use this as a leading candidate, then validate the README and install path in your own agent stack.
Fastest prototype
skill-upper
Best first install candidate based on install readiness and adoption.
Freshest repo
skill-upper
Most recent maintenance signal among this shortlist.
| Signal | skill-upper Create, run, diagnose, and iteratively improve Agent Skill evaluations (evals) with the skill-up CLI / 使用 skill-up CLI 创建、运行、诊断并持续改进 Agent Skill 评测. Use when the user asks to evaluate, test, regress, verify, fix, improve, iterate, or evolve a Skill; add or strengthen eval cases; write eval.yaml/case.yaml; run skill-up run/validate/list-cases/report/import/init; or migrate from Anthropic evals.json. Handles Skill discovery, eval scaffolding, judge authoring, validation, runs, reports, and evidence-based repair loops. |
|---|---|
| Quality | 76/100 Strong |
| Decision verdict | 87/100 Production-ready Use this as a leading candidate, then validate the README and install path in your own agent stack. |
| Adoption | 846 stars Verified outcomes are shown on each skill page |
| Freshness | Sep 4, 2026 |
| Use-case fit | |
| Workflow fit | |
| Platform hints | Claude Code, OpenAI Agents |
| Warnings | No OpenAgentSkill engagement data yet |
| Best for | Research agents workflows · Claude Code teams · teams that value GitHub adoption signals |
Skill comparison
Use this as a shortlist, then open the skill detail page before adopting.
Decision summary
Strongest overall
skill-upper
Use this as a leading candidate, then validate the README and install path in your own agent stack.
Fastest prototype
skill-upper
Best first install candidate based on install readiness and adoption.
Freshest repo
skill-upper
Most recent maintenance signal among this shortlist.
| Signal | skill-upper Create, run, diagnose, and iteratively improve Agent Skill evaluations (evals) with the skill-up CLI / 使用 skill-up CLI 创建、运行、诊断并持续改进 Agent Skill 评测. Use when the user asks to evaluate, test, regress, verify, fix, improve, iterate, or evolve a Skill; add or strengthen eval cases; write eval.yaml/case.yaml; run skill-up run/validate/list-cases/report/import/init; or migrate from Anthropic evals.json. Handles Skill discovery, eval scaffolding, judge authoring, validation, runs, reports, and evidence-based repair loops. |
|---|---|
| Quality | 76/100 Strong |
| Decision verdict | 87/100 Production-ready Use this as a leading candidate, then validate the README and install path in your own agent stack. |
| Adoption | 846 stars Verified outcomes are shown on each skill page |
| Freshness | Sep 4, 2026 |
| Use-case fit | |
| Workflow fit | |
| Platform hints | Claude Code, OpenAI Agents |
| Warnings | No OpenAgentSkill engagement data yet |
| Best for | Research agents workflows · Claude Code teams · teams that value GitHub adoption signals |
| Not ideal for | teams that need a vendor-supported SLA · high-compliance environments without internal security review |
| OpenAgentSkill engagement | 0 views 0 install copies |
| Install | $ npx skills add alibaba/skill-up --skill skill-upper |
| Not ideal for | teams that need a vendor-supported SLA · high-compliance environments without internal security review |
| OpenAgentSkill engagement | 0 views 0 install copies |
| Install | $ npx skills add alibaba/skill-up --skill skill-upper |