Skill audit report
BenchMARL Audit report.
BenchMARL is a library for benchmarking Multi-Agent Reinforcement Learning (MARL). BenchMARL allows to quickly compare different MARL algorithms, tasks, and models while being systematically grounded in its two core tenets: reproducibility and standardization.
OpenAgentSkill Trust Score
OpenAgentSkill Trust Score
The Trust Score helps an agent decide whether a skill is safe enough to shortlist before installation.
GitHub adoption
INFO76
636 GitHub stars
Stars/forks activity
INFO71
636 stars, 132 forks; issue activity unavailable in current metadata
Recent maintenance
INFO76
6mo since push
License clarity
PASS86
MIT
README/SKILL.md completeness
PASS90
Metadata includes enough usage and workflow context
Dependency/runtime risk
PASS90
no major dependency risk hints in public metadata
Install availability
PASS92
npx skills add facebookresearch/BenchMARL
Install command safety
PASS92
standard package or runtime install path
Permission surface
PASS86
filesystem or document access
Repository evidence
PASS86
https://github.com/facebookresearch/BenchMARL
Review status
PASS88
AI review data available
Agent Proven outcomes
INFO54
No agent outcome data yet
Checks
Install and adoption review
Install path
92
npx skills add facebookresearch/BenchMARL
Repository
88
https://github.com/facebookresearch/BenchMARL
License
86
MIT
Maintenance
76
6mo since push
AI review
88
Approved with no listed issues
README/SKILL.md completeness
90
Usable description available
Dependency risk
90
no major dependency risk hints in public metadata
Install command safety
92
standard package or runtime install path
Permission surface
86
filesystem or document access
Stars/forks activity
71
636 stars, 132 forks; issue activity unavailable in current metadata
Adoption
88
636 GitHub stars
Warnings
- Quality score needs review
Method
This report combines public metadata, AI review output, repository freshness, install readiness, OpenAgentSkill events, quality scoring, trust checks, and the agent safety gate. It is not a full source-code security review.