Skill audit report
Use this skill whenever the user wants to evaluate, test, or validate an AI agent, decide whether an agent is ready to ship or go live, choose how to grade an agent's answers (exact match, similarity, meaning, keywords, quality, or custom), design a test set of questions and expected answers, or interpret evaluation results into a go/no-go decision. Invoke it before the user hand-builds tests or declares an agent "done."
OpenAgentSkill Trust Score
The Trust Score helps an agent decide whether a skill is safe enough to shortlist before installation.
GitHub adoption
WARN48
66 GitHub stars
Stars/forks activity
WARN53
66 stars, 88 forks; issue activity unavailable in current metadata
Recent maintenance
PASS100
24d since push
License clarity
PASS86
MIT
README/SKILL.md completeness
PASS86
Metadata includes enough usage and workflow context
Dependency/runtime risk
INFO72
credential or environment access
Install availability
PASS92
npx skills add microsoft/cat-agent-skills --skill agent-evaluation-designer
Install command safety
PASS92
standard package or runtime install path
Permission surface
INFO74
secrets or environment access
Repository evidence
PASS86
https://github.com/microsoft/cat-agent-skills/tree/main/submissions/agent-evaluation-designer
Review status
WARN46
AI review approval is missing
Agent Proven outcomes
INFO54
No agent outcome data yet
Checks
Install path
92
npx skills add microsoft/cat-agent-skills --skill agent-evaluation-designer
Repository
88
https://github.com/microsoft/cat-agent-skills/tree/main/submissions/agent-evaluation-designer
License
86
MIT
Maintenance
100
24d since push
AI review
55
Review approval is missing
README/SKILL.md completeness
86
Usable description available
Dependency risk
72
credential or environment access
Install command safety
92
standard package or runtime install path
Permission surface
74
secrets or environment access
Stars/forks activity
53
66 stars, 88 forks; issue activity unavailable in current metadata
Adoption
68
66 GitHub stars
Warnings
Method
This report combines public metadata, AI review output, repository freshness, install readiness, OpenAgentSkill events, quality scoring, trust checks, and the agent safety gate. It is not a full source-code security review.
Compare nearby options
Guidance for distinctive, intentional UI design, typography, visual direction, and non-template-like product interfaces.
180K Stars · Audit report
Design and implementation guidance for distinctive landing pages, portfolios, product demos, and purposeful redesigns.
92K Stars · Audit report
Create original visual art, posters, PNG assets, and PDF documents through a clear design philosophy.
180K Stars · Audit report