Skill audit report
Use this to measure whether an AI agent actually completed its task end to end, not just whether individual LLM calls looked fine. Trigger on "is my agent working", "measure agent success rate", "evaluate my agent", "how good is my agent", "agent completion rate", or evaluating a multi-step/tool-using agent. Score the outcome of the whole task, plus the path it took.
OpenAgentSkill Trust Score
The Trust Score helps an agent decide whether a skill is safe enough to shortlist before installation.
GitHub adoption
WARN48
33 GitHub stars
Stars/forks activity
WARN48
33 stars, 18 forks; issue activity unavailable in current metadata
Recent maintenance
PASS100
10d since push
License clarity
PASS86
CC0-1.0
README/SKILL.md completeness
INFO76
Public metadata needs stronger README/SKILL.md context
Dependency/runtime risk
PASS82
network or browser surface
Install availability
PASS92
npx skills add ContextJet-ai/awesome-llm-observability --skill measure-agent-task-success
Install command safety
PASS92
standard package or runtime install path
Permission surface
PASS86
network or browser access
Repository evidence
PASS86
https://github.com/ContextJet-ai/awesome-llm-observability/tree/main/skills/measure-agent-task-success
Review status
WARN46
AI review approval is missing
Agent Proven outcomes
INFO54
No agent outcome data yet
Checks
Install path
92
npx skills add ContextJet-ai/awesome-llm-observability --skill measure-agent-task-success
Repository
88
https://github.com/ContextJet-ai/awesome-llm-observability/tree/main/skills/measure-agent-task-success
License
86
CC0-1.0
Maintenance
100
10d since push
AI review
55
Review approval is missing
README/SKILL.md completeness
84
Usable description available
Dependency risk
82
network or browser surface
Install command safety
92
standard package or runtime install path
Permission surface
86
network or browser access
Stars/forks activity
48
33 stars, 18 forks; issue activity unavailable in current metadata
Adoption
42
33 GitHub stars
Warnings
Method
This report combines public metadata, AI review output, repository freshness, install readiness, OpenAgentSkill events, quality scoring, trust checks, and the agent safety gate. It is not a full source-code security review.
Compare nearby options