Skill audit report
This skill should be used when the user asks to "create evals", "evaluate an agent", "build evaluation suite", or mentions agent testing, graders, or benchmarks. Also suggest when building coding agents, conversational agents, or research agents that need quality assurance.
The Trust Score helps an agent decide whether a skill is safe enough to shortlist before installation.
GitHub adoption
FAIL30
23 GitHub stars
Stars/forks activity
FAIL32
23 stars, 2 forks; issue activity unavailable in current metadata
Recent maintenance
PASS100
7d since push
License clarity
PASS86
MIT
README/SKILL.md completeness
PASS86
Metadata includes enough usage and workflow context
Dependency/runtime risk
INFO72
credential or environment access
Install availability
PASS92
npx skills add dwmkerr/claude-toolkit --skill anthropic-evaluations
Install command safety
PASS92
standard package or runtime install path
Permission surface
INFO74
secrets or environment access
Repository evidence
PASS86
https://github.com/dwmkerr/claude-toolkit/tree/main/plugins/toolkit/skills/anthropic-evaluations
Review status
WARN46
AI review approval is missing
Agent Proven outcomes
INFO54
No agent outcome data yet
Checks
Install path
92
npx skills add dwmkerr/claude-toolkit --skill anthropic-evaluations
Repository
88
https://github.com/dwmkerr/claude-toolkit/tree/main/plugins/toolkit/skills/anthropic-evaluations
License
86
MIT
Maintenance
100
7d since push
AI review
55
Review approval is missing
README/SKILL.md completeness
86
Usable description available
Dependency risk
72
credential or environment access
Install command safety
92
standard package or runtime install path
Permission surface
74
secrets or environment access
Stars/forks activity
32
23 stars, 2 forks; issue activity unavailable in current metadata
Adoption
42
23 GitHub stars
Warnings
Method
This report combines public metadata, AI review output, repository freshness, install readiness, OpenAgentSkill events, quality scoring, trust checks, and the agent safety gate. It is not a full source-code security review.
Compare nearby options
Generate original one-ink or controlled two-ink editorial images from any theme, sentence, article idea, object, or reference photo. Always use this skill when the user asks for 单色海报、双色印刷、单色调视觉、蓝色/绿色孔版印刷、risograph、网点照片、复古或当代编辑排版、zine poster, monochrome editorial poster, duotone print, or asks to use the mono-color style. It uses an adaptive white, gray, or pale-beige substrate, no more than two printing inks, active negative space, terse human language, and strong serif/grotesk/mono typography without making retro styling the default or copying a source composition, wording, logo, or artwork. Produce both the final generation prompt and the generated raster image unless the user explicitly asks for prompt only.
1.9K Stars · Audit report
Research the last 30 days across Reddit, X, YouTube, Hacker News, Polymarket, GitHub, and the web, then synthesize a grounded brief for an AI agent.
62K Stars · Audit report
Academic Research Skills for Claude Code: research → write → review → revise → finalize
38K Stars · Audit report