Skill audit report

AgentEval Audit report.

AgentEval is the comprehensive .NET toolkit for AI agent evaluation—tool usage validation, RAG quality metrics, stochastic evaluation, and model comparison—built first for Microsoft Agent Framework (MAF) and Microsoft.Extensions.AI. What RAGAS, PromptFoo and DeepEval do for Python, AgentEval does for .NET

REVIEWED · REVIEWNeeds reviewGenerated Aug 23, 2026Heuristic metadata audit
86
Audit
83
Trust
78
Quality
89
Security
100
Maintain
92
Install

OpenAgentSkill Trust Score

83
Strong shortlist

OpenAgentSkill Trust Score

The Trust Score helps an agent decide whether a skill is safe enough to shortlist before installation.

GitHub adoption

INFO

62

133 GitHub stars

Stars/forks activity

WARN

57

133 stars, 12 forks; issue activity unavailable in current metadata

Recent maintenance

PASS

100

16d since push

License clarity

PASS

86

MIT

README/SKILL.md completeness

PASS

100

Metadata includes enough usage and workflow context

Dependency/runtime risk

PASS

90

no major dependency risk hints in public metadata

Install availability

PASS

92

npx skills add AgentEvalHQ/AgentEval

Install command safety

PASS

92

standard package or runtime install path

Permission surface

PASS

86

filesystem or document access

Repository evidence

PASS

86

https://github.com/AgentEvalHQ/AgentEval

Review status

PASS

88

AI review data available

Agent Proven outcomes

INFO

54

No agent outcome data yet

Checks

Install and adoption review

9 Passed · 6 Needs review

Install path

92

PASS

npx skills add AgentEvalHQ/AgentEval

Repository

88

PASS

https://github.com/AgentEvalHQ/AgentEval

License

86

PASS

MIT

Maintenance

100

PASS

16d since push

AI review

88

PASS

Approved with no listed issues

README/SKILL.md completeness

100

PASS

Usable description available

Dependency risk

90

PASS

no major dependency risk hints in public metadata

Install command safety

92

PASS

standard package or runtime install path

Permission surface

86

PASS

filesystem or document access

Stars/forks activity

57

FIX

133 stars, 12 forks; issue activity unavailable in current metadata

Adoption

68

INFO

133 GitHub stars

Financial decision safety

58

CHECK

Research-only use: do not treat output as financial advice or execute a position without human approval.

Warnings

  • Financial research output is not financial advice; require human review before any live investment decision
  • Financial research output is not financial advice; require human review before any live investment decision.
  • Quality score needs review
  • Stars/forks activity: 133 stars, 12 forks; issue activity unavailable in current metadata

Method

This report combines public metadata, AI review output, repository freshness, install readiness, OpenAgentSkill events, quality scoring, trust checks, and the agent safety gate. It is not a full source-code security review.

Compare nearby options

Related skills to audit next