Skill audit report
Design and validity review for studies that benchmark one or more AI systems against a human-expert panel as the reference. Covers the evaluation question and arm definition, decoupled multi-dimensional rubrics with anchors, planted calibration probes, reviewer-panel construction, inter-rater reliability targets, LLM-as-judge versus human-as-judge adjudication, construct-independence guards, and a structured rating-export schema. Use before data collection on an AI-vs-expert evaluation.
Skill audit report
Design and validity review for studies that benchmark one or more AI systems against a human-expert panel as the reference. Covers the evaluation question and arm definition, decoupled multi-dimensional rubrics with anchors, planted calibration probes, reviewer-panel construction, inter-rater reliability targets, LLM-as-judge versus human-as-judge adjudication, construct-independence guards, and a structured rating-export schema. Use before data collection on an AI-vs-expert evaluation.
Skill audit report
Design and validity review for studies that benchmark one or more AI systems against a human-expert panel as the reference. Covers the evaluation question and arm definition, decoupled multi-dimensional rubrics with anchors, planted calibration probes, reviewer-panel construction, inter-rater reliability targets, LLM-as-judge versus human-as-judge adjudication, construct-independence guards, and a structured rating-export schema. Use before data collection on an AI-vs-expert evaluation.
Skill audit report
Design and validity review for studies that benchmark one or more AI systems against a human-expert panel as the reference. Covers the evaluation question and arm definition, decoupled multi-dimensional rubrics with anchors, planted calibration probes, reviewer-panel construction, inter-rater reliability targets, LLM-as-judge versus human-as-judge adjudication, construct-independence guards, and a structured rating-export schema. Use before data collection on an AI-vs-expert evaluation.
OpenAgentSkill Trust Score
The Trust Score helps an agent decide whether a skill is safe enough to shortlist before installation.
GitHub adoption
INFO62
283 GitHub stars
Stars/forks activity
INFO62
283 stars, 69 forks; issue activity unavailable in current metadata
Recent maintenance
PASS100
Pushed today
License clarity
PASS86
MIT
README/SKILL.md completeness
INFO76
Public metadata needs stronger README/SKILL.md context
Dependency/runtime risk
INFO72
command execution surface
Install availability
PASS92
npx skills add Aperivue/medsci-skills --skill design-ai-benchmarking
Install command safety
PASS92
standard package or runtime install path
Permission surface
INFO64
shell or command execution, database access
Repository evidence
PASS86
https://github.com/Aperivue/medsci-skills/tree/main/skills/design-ai-benchmarking
Review status
INFO66
AI review data available
Agent Proven outcomes
INFO54
No agent outcome data yet
Checks
Install path
92
npx skills add Aperivue/medsci-skills --skill design-ai-benchmarking
Repository
88
https://github.com/Aperivue/medsci-skills/tree/main/skills/design-ai-benchmarking
License
86
MIT
Maintenance
100
Pushed today
AI review
55
No critical security issues detected; the skill is advisory and uses standard file tools only.
README/SKILL.md completeness
84
Usable description available
Dependency risk
72
command execution surface
Install command safety
92
standard package or runtime install path
Permission surface
64
shell or command execution, database access
Stars/forks activity
62
283 stars, 69 forks; issue activity unavailable in current metadata
Adoption
68
283 GitHub stars
Warnings
Method
This report combines public metadata, AI review output, repository freshness, install readiness, OpenAgentSkill events, quality scoring, trust checks, and the agent safety gate. It is not a full source-code security review.
Compare nearby options
Guidance for distinctive, intentional UI design, typography, visual direction, and non-template-like product interfaces.
175K Stars · Audit report
Design and implementation guidance for distinctive landing pages, portfolios, product demos, and purposeful redesigns.
85K Stars · Audit report
Turn one topic into a narrated Vox-style paper-collage explainer or ad video, from script through captions.
1.8K Stars · Audit report
OpenAgentSkill Trust Score
The Trust Score helps an agent decide whether a skill is safe enough to shortlist before installation.
GitHub adoption
INFO62
283 GitHub stars
Stars/forks activity
INFO62
283 stars, 69 forks; issue activity unavailable in current metadata
Recent maintenance
PASS100
Pushed today
License clarity
PASS86
MIT
README/SKILL.md completeness
INFO76
Public metadata needs stronger README/SKILL.md context
Dependency/runtime risk
INFO72
command execution surface
Install availability
PASS92
npx skills add Aperivue/medsci-skills --skill design-ai-benchmarking
Install command safety
PASS92
standard package or runtime install path
Permission surface
INFO64
shell or command execution, database access
Repository evidence
PASS86
https://github.com/Aperivue/medsci-skills/tree/main/skills/design-ai-benchmarking
Review status
INFO66
AI review data available
Agent Proven outcomes
INFO54
No agent outcome data yet
Checks
Install path
92
npx skills add Aperivue/medsci-skills --skill design-ai-benchmarking
Repository
88
https://github.com/Aperivue/medsci-skills/tree/main/skills/design-ai-benchmarking
License
86
MIT
Maintenance
100
Pushed today
AI review
55
No critical security issues detected; the skill is advisory and uses standard file tools only.
README/SKILL.md completeness
84
Usable description available
Dependency risk
72
command execution surface
Install command safety
92
standard package or runtime install path
Permission surface
64
shell or command execution, database access
Stars/forks activity
62
283 stars, 69 forks; issue activity unavailable in current metadata
Adoption
68
283 GitHub stars
Warnings
Method
This report combines public metadata, AI review output, repository freshness, install readiness, OpenAgentSkill events, quality scoring, trust checks, and the agent safety gate. It is not a full source-code security review.
Compare nearby options
Guidance for distinctive, intentional UI design, typography, visual direction, and non-template-like product interfaces.
175K Stars · Audit report
Design and implementation guidance for distinctive landing pages, portfolios, product demos, and purposeful redesigns.
85K Stars · Audit report
Turn one topic into a narrated Vox-style paper-collage explainer or ad video, from script through captions.
1.8K Stars · Audit report
OpenAgentSkill Trust Score
The Trust Score helps an agent decide whether a skill is safe enough to shortlist before installation.
GitHub adoption
INFO62
283 GitHub stars
Stars/forks activity
INFO62
283 stars, 69 forks; issue activity unavailable in current metadata
Recent maintenance
PASS100
Pushed today
License clarity
PASS86
MIT
README/SKILL.md completeness
INFO76
Public metadata needs stronger README/SKILL.md context
Dependency/runtime risk
INFO72
command execution surface
Install availability
PASS92
npx skills add Aperivue/medsci-skills --skill design-ai-benchmarking
Install command safety
PASS92
standard package or runtime install path
Permission surface
INFO64
shell or command execution, database access
Repository evidence
PASS86
https://github.com/Aperivue/medsci-skills/tree/main/skills/design-ai-benchmarking
Review status
INFO66
AI review data available
Agent Proven outcomes
INFO54
No agent outcome data yet
Checks
Install path
92
npx skills add Aperivue/medsci-skills --skill design-ai-benchmarking
Repository
88
https://github.com/Aperivue/medsci-skills/tree/main/skills/design-ai-benchmarking
License
86
MIT
Maintenance
100
Pushed today
AI review
55
No critical security issues detected; the skill is advisory and uses standard file tools only.
README/SKILL.md completeness
84
Usable description available
Dependency risk
72
command execution surface
Install command safety
92
standard package or runtime install path
Permission surface
64
shell or command execution, database access
Stars/forks activity
62
283 stars, 69 forks; issue activity unavailable in current metadata
Adoption
68
283 GitHub stars
Warnings
Method
This report combines public metadata, AI review output, repository freshness, install readiness, OpenAgentSkill events, quality scoring, trust checks, and the agent safety gate. It is not a full source-code security review.
Compare nearby options
Guidance for distinctive, intentional UI design, typography, visual direction, and non-template-like product interfaces.
175K Stars · Audit report
Design and implementation guidance for distinctive landing pages, portfolios, product demos, and purposeful redesigns.
85K Stars · Audit report
Turn one topic into a narrated Vox-style paper-collage explainer or ad video, from script through captions.
1.8K Stars · Audit report
OpenAgentSkill Trust Score
The Trust Score helps an agent decide whether a skill is safe enough to shortlist before installation.
GitHub adoption
INFO62
283 GitHub stars
Stars/forks activity
INFO62
283 stars, 69 forks; issue activity unavailable in current metadata
Recent maintenance
PASS100
Pushed today
License clarity
PASS86
MIT
README/SKILL.md completeness
INFO76
Public metadata needs stronger README/SKILL.md context
Dependency/runtime risk
INFO72
command execution surface
Install availability
PASS92
npx skills add Aperivue/medsci-skills --skill design-ai-benchmarking
Install command safety
PASS92
standard package or runtime install path
Permission surface
INFO64
shell or command execution, database access
Repository evidence
PASS86
https://github.com/Aperivue/medsci-skills/tree/main/skills/design-ai-benchmarking
Review status
INFO66
AI review data available
Agent Proven outcomes
INFO54
No agent outcome data yet
Checks
Install path
92
npx skills add Aperivue/medsci-skills --skill design-ai-benchmarking
Repository
88
https://github.com/Aperivue/medsci-skills/tree/main/skills/design-ai-benchmarking
License
86
MIT
Maintenance
100
Pushed today
AI review
55
No critical security issues detected; the skill is advisory and uses standard file tools only.
README/SKILL.md completeness
84
Usable description available
Dependency risk
72
command execution surface
Install command safety
92
standard package or runtime install path
Permission surface
64
shell or command execution, database access
Stars/forks activity
62
283 stars, 69 forks; issue activity unavailable in current metadata
Adoption
68
283 GitHub stars
Warnings
Method
This report combines public metadata, AI review output, repository freshness, install readiness, OpenAgentSkill events, quality scoring, trust checks, and the agent safety gate. It is not a full source-code security review.
Compare nearby options
Guidance for distinctive, intentional UI design, typography, visual direction, and non-template-like product interfaces.
175K Stars · Audit report
Design and implementation guidance for distinctive landing pages, portfolios, product demos, and purposeful redesigns.
85K Stars · Audit report
Turn one topic into a narrated Vox-style paper-collage explainer or ad video, from script through captions.
1.8K Stars · Audit report