Skill comparison
Compare agent skills before installing.
Comparing 1 skill
Use this as a shortlist, then open the skill detail page before adopting.
Decision summary
arbor is the strongest overall pick here because it has a 100/100 readiness score and fits Research agents.
Strongest overall
arbor
Use this as a leading candidate, then validate the README and install path in your own agent stack.
Fastest prototype
arbor
Best first install candidate based on install readiness and adoption.
Freshest repo
arbor
Most recent maintenance signal among this shortlist.
| Signal | arbor Autonomously improve a real artifact (code, training recipe, agent harness, data pipeline, prompt) against an objective and an evaluator, using Hypothesis Tree Refinement (HTR) from the Arbor paper. Use this whenever someone wants to iteratively optimize something over many experiments without overfitting — e.g. "get my model's eval score up", "improve this agent/harness", "tune this pipeline", "beat the baseline on this benchmark", "run a search over approaches and keep the best", "do an MLE-bench / Kaggle-style optimization", or any long-horizon "make this artifact better and don't just memorize the dev set" task. Trigger it even when the user doesn't say "Arbor" or "hypothesis tree" but describes repeated experiment-and-evaluate loops, branching exploration of competing ideas, or worries about a dev/test gap. Runs Claude itself as the coordinator with subagent executors in isolated git worktrees; for the standalone `arbor` CLI tool see references/arbor-upstream.md. |
|---|---|
| Quality | 92/100 Excellent |
| Decision verdict | 100/100 Production-ready Use this as a leading candidate, then validate the README and install path in your own agent stack. |
| Adoption | 34K stars Verified outcomes are shown on each skill page |
| Freshness | Aug 20, 2026 |
| Use-case fit | |
| Workflow fit | |
| Platform hints | Claude Code |
| Warnings | No OpenAgentSkill engagement data yet |
| Best for | Research agents workflows · Claude Code teams · teams that value GitHub adoption signals |
| Not ideal for | teams that need a vendor-supported SLA · high-compliance environments without internal security review |
| OpenAgentSkill engagement | 0 views 0 install copies |
| Install | $ npx skills add K-Dense-AI/scientific-agent-skills --skill arbor |