Skill comparison
Use this as a shortlist, then open the skill detail page before adopting.
Decision summary
Strongest overall
VLMEvalKit
Use this as a leading candidate, then validate the README and install path in your own agent stack.
Fastest prototype
VLMEvalKit
Best first install candidate based on install readiness and adoption.
Freshest repo
Semantic Router
Most recent maintenance signal among this shortlist.
| Signal | Benchbot BenchBot is a tool for seamlessly testing & evaluating semantic scene understanding tools in both realistic 3D simulation & on real robots | VLMEvalKit Open-source evaluation toolkit of large multi-modality models (LMMs), support 220+ LMMs, 80+ benchmarks | Semantic Router Superfast AI decision making and intelligent processing of multi-modal data. | Aily Blockly AI IDE for hardware development, support Arduino, MicroPython, ESP32, STM32, RP2040, Nrf5x... |
|---|---|---|---|---|
| Quality | 47/100 Needs review | 100/100 Excellent | 100/100 Excellent | 100/100 Excellent |
| Decision verdict | 37/100 Needs manual review Do a manual repository review before adding this to an agent workflow. | 100/100 Production-ready Use this as a leading candidate, then validate the README and install path in your own agent stack. | 100/100 Production-ready Use this as a leading candidate, then validate the README and install path in your own agent stack. | 100/100 Production-ready Use this as a leading candidate, then validate the README and install path in your own agent stack. |
| Adoption | 113 stars Verified outcomes are shown on each skill page | 4.2K stars Verified outcomes are shown on each skill page | 3.9K stars Verified outcomes are shown on each skill page |
| 3.8K stars Verified outcomes are shown on each skill page |
| Freshness | Aug 13, 2023 | Jun 17, 2026 | Sep 12, 2026 | Aug 11, 2026 |
| Use-case fit |
| Workflow fit |
| Platform hints | Shell, Robotics, Claude Code | Python, Computer Vision, Claude Code, OpenAI Agents | Python, Computer Vision, Claude Code | TypeScript, IoT, Claude Code |
| Warnings | Repository looks stale · No OpenAgentSkill engagement data yet | No OpenAgentSkill engagement data yet | No OpenAgentSkill engagement data yet | No OpenAgentSkill engagement data yet |
| Best for | Coding agents workflows · Claude Code teams · builders willing to evaluate younger projects | Browser automation workflows · Claude Code teams · teams that value GitHub adoption signals | RAG and knowledge workflows · Claude Code teams · teams that value GitHub adoption signals | Coding agents workflows · Claude Code teams · teams that value GitHub adoption signals |
| Not ideal for | teams that require actively maintained dependencies · production agents without a repository review | teams that need a vendor-supported SLA · high-compliance environments without internal security review | teams that need a vendor-supported SLA · high-compliance environments without internal security review | teams that need a vendor-supported SLA · high-compliance environments without internal security review |
| OpenAgentSkill engagement | 0 views 0 install copies | 0 views 0 install copies | 0 views 0 install copies | 0 views 0 install copies |
| Install | $ npx skills add qcr/benchbot | $ npx skills add open-compass/VLMEvalKit | $ npx skills add aurelio-labs/semantic-router | $ npx skills add ailyProject/aily-blockly |