Skill comparison
Use this as a shortlist, then open the skill detail page before adopting.
Decision summary
Strongest overall
Evalscope
Use this as a leading candidate, then validate the README and install path in your own agent stack.
Fastest prototype
Evalscope
Best first install candidate based on install readiness and adoption.
Freshest repo
Evalscope
Most recent maintenance signal among this shortlist.
| Signal | Evalscope A streamlined and customizable framework for efficient large model (LLM, VLM, AIGC) evaluation and performance benchmarking. |
|---|---|
| Quality | 100/100 Excellent |
| Decision verdict | 100/100 Production-ready Use this as a leading candidate, then validate the README and install path in your own agent stack. |
| Adoption | 3.0K stars 0 installs |
| Freshness | Jun 18, 2026 |
| Use-case fit | |
| Stack fit | |
| Platform hints | Python, RAG, Claude Code |
| Warnings | No major risk signals from current metadata |
| Best for | RAG and knowledge workflows · Claude Code teams · teams that value GitHub adoption signals |
| Not ideal for | teams that need a vendor-supported SLA · high-compliance environments without internal security review |
| OpenAgentSkill engagement | 8 views 0 install copies |
| Install | $ npx skills add modelscope/evalscope |