Skill comparison
Use this as a shortlist, then open the skill detail page before adopting.
Decision summary
Strongest overall
Code Review
Use this as a leading candidate, then validate the README and install path in your own agent stack.
Fastest prototype
Code Review
Best first install candidate based on install readiness and adoption.
Freshest repo
evaluate-run
Most recent maintenance signal among this shortlist.
| Signal | evaluate-run Grade a completed factory run (a worktree + its chat) against your standards, cold and adversarially, so the system learns. Squashes the session transcript into a compact digest (prompts + corrections, the ordered tool timeline, the diff), then a fresh evaluator judges how the run went - did it follow the loop, honor your conventions, write real tests, ask at the right moments, handle corrections - and routes every finding to a durable fix (/add-rule for code, a persistent memory for process). This is the OFFLINE complement to the correction hook - the hook catches what you notice live, this catches what you didn't. Use for "evaluate this factory run", "grade this chat", "how did that go, what should it learn". | Code Review Review a branch or diff against repository standards and the originating spec in two independent analysis passes. |
|---|---|---|
| Quality | 55/100 Promising | 100/100 Excellent |
| Decision verdict | 54/100 Needs manual review Do a manual repository review before adding this to an agent workflow. | 100/100 Production-ready Use this as a leading candidate, then validate the README and install path in your own agent stack. |
| Adoption | 22 stars Verified outcomes are shown on each skill page | 169K stars Verified outcomes are shown on each skill page |
| Freshness | Sep 4, 2026 | Jul 13, 2026 |
| Use-case fit |
| Workflow fit |
| Platform hints | Claude Code | Claude Code, Codex, Cursor, OpenAI Agents |
| Warnings | Low GitHub adoption signal · No OpenAgentSkill engagement data yet | No OpenAgentSkill engagement data yet |
| Best for | Coding agents workflows · Claude Code teams · builders willing to evaluate younger projects | Coding agents workflows · Claude Code teams · teams that value GitHub adoption signals |
| Not ideal for | teams that need a vendor-supported SLA · production agents without a repository review | teams that need a vendor-supported SLA · high-compliance environments without internal security review |
| OpenAgentSkill engagement | 0 views 0 install copies | 0 views 0 install copies |
| Install | $ npx skills add uiverify/uiverify --skill evaluate-run | $ npx skills add mattpocock/skills --skill code-review |