Skill audit report
Framework for measuring and tracking agent response quality over time. Detects regressions before they reach production. Use when evaluating agent changes, auditing quality, or establishing performance baselines.
OpenAgentSkill Trust Score
The Trust Score helps an agent decide whether a skill is safe enough to shortlist before installation.
GitHub adoption
INFO76
530 GitHub stars
Stars/forks activity
INFO65
530 stars, 44 forks; issue activity unavailable in current metadata
Recent maintenance
PASS88
1mo since push
License clarity
PASS86
MIT
README/SKILL.md completeness
PASS86
Metadata includes enough usage and workflow context
Dependency/runtime risk
FAIL38
command execution surface, credential or environment access
Install availability
PASS92
npx skills add vibeeval/vibecosystem --skill agent-benchmark
Install command safety
PASS92
standard package or runtime install path
Permission surface
FAIL18
secrets or environment access, shell or command execution
Repository evidence
PASS86
https://github.com/vibeeval/vibecosystem/tree/main/skills/agent-benchmark
Review status
INFO66
AI review data available
Agent Proven outcomes
INFO54
No agent outcome data yet
Checks
Install path
92
npx skills add vibeeval/vibecosystem --skill agent-benchmark
Repository
88
https://github.com/vibeeval/vibecosystem/tree/main/skills/agent-benchmark
License
86
MIT
Maintenance
88
1mo since push
AI review
55
The skill references a run script (run.mjs) and directory structure but does not include installation or setup instructions.
README/SKILL.md completeness
86
Usable description available
Dependency risk
38
command execution surface, credential or environment access
Install command safety
92
standard package or runtime install path
Permission surface
18
secrets or environment access, shell or command execution
Stars/forks activity
65
530 stars, 44 forks; issue activity unavailable in current metadata
Adoption
88
530 GitHub stars
Warnings
Method
This report combines public metadata, AI review output, repository freshness, install readiness, OpenAgentSkill events, quality scoring, trust checks, and the agent safety gate. It is not a full source-code security review.