Skill audit report
End-to-end Agent Observability pipeline for an instrumented ml_app — classify production traces, root-cause failures, bootstrap evaluators, then (optionally) sample + publish a dataset, generate + run an experiment, and analyze results. Six narrated phases with a standardized banner and a "continue" checkpoint between each. Pure orchestration over the agent-observability sub-skills (`agent-observability-session-classify`, `agent-observability-trace-rca`, `agent-observability-eval-bootstrap`, `agent-observability-experiment-bootstrap`, `agent-observability-experiment-analyzer`). Use when user says "run the eval pipeline", "go from traces to evals", "bootstrap evals end to end", "classify then RCA then bootstrap", "build an eval set from scratch", "onboard me to datasets and experiments", "walk me through experiments", "I have an ml_app, now what", "Agent Observability onboarding", "guided experiment setup", "from traces to experiments", or wants a deterministic, narrated tour from produ
Skill audit report
End-to-end Agent Observability pipeline for an instrumented ml_app — classify production traces, root-cause failures, bootstrap evaluators, then (optionally) sample + publish a dataset, generate + run an experiment, and analyze results. Six narrated phases with a standardized banner and a "continue" checkpoint between each. Pure orchestration over the agent-observability sub-skills (`agent-observability-session-classify`, `agent-observability-trace-rca`, `agent-observability-eval-bootstrap`, `agent-observability-experiment-bootstrap`, `agent-observability-experiment-analyzer`). Use when user says "run the eval pipeline", "go from traces to evals", "bootstrap evals end to end", "classify then RCA then bootstrap", "build an eval set from scratch", "onboard me to datasets and experiments", "walk me through experiments", "I have an ml_app, now what", "Agent Observability onboarding", "guided experiment setup", "from traces to experiments", or wants a deterministic, narrated tour from produ
Skill audit report
End-to-end Agent Observability pipeline for an instrumented ml_app — classify production traces, root-cause failures, bootstrap evaluators, then (optionally) sample + publish a dataset, generate + run an experiment, and analyze results. Six narrated phases with a standardized banner and a "continue" checkpoint between each. Pure orchestration over the agent-observability sub-skills (`agent-observability-session-classify`, `agent-observability-trace-rca`, `agent-observability-eval-bootstrap`, `agent-observability-experiment-bootstrap`, `agent-observability-experiment-analyzer`). Use when user says "run the eval pipeline", "go from traces to evals", "bootstrap evals end to end", "classify then RCA then bootstrap", "build an eval set from scratch", "onboard me to datasets and experiments", "walk me through experiments", "I have an ml_app, now what", "Agent Observability onboarding", "guided experiment setup", "from traces to experiments", or wants a deterministic, narrated tour from produ
Skill audit report
End-to-end Agent Observability pipeline for an instrumented ml_app — classify production traces, root-cause failures, bootstrap evaluators, then (optionally) sample + publish a dataset, generate + run an experiment, and analyze results. Six narrated phases with a standardized banner and a "continue" checkpoint between each. Pure orchestration over the agent-observability sub-skills (`agent-observability-session-classify`, `agent-observability-trace-rca`, `agent-observability-eval-bootstrap`, `agent-observability-experiment-bootstrap`, `agent-observability-experiment-analyzer`). Use when user says "run the eval pipeline", "go from traces to evals", "bootstrap evals end to end", "classify then RCA then bootstrap", "build an eval set from scratch", "onboard me to datasets and experiments", "walk me through experiments", "I have an ml_app, now what", "Agent Observability onboarding", "guided experiment setup", "from traces to experiments", or wants a deterministic, narrated tour from produ
OpenAgentSkill Trust Score
The Trust Score helps an agent decide whether a skill is safe enough to shortlist before installation.
GitHub adoption
INFO62
158 GitHub stars
Stars/forks activity
WARN57
158 stars, 25 forks; issue activity unavailable in current metadata
Recent maintenance
PASS100
9d since push
License clarity
PASS86
MIT
README/SKILL.md completeness
PASS86
Metadata includes enough usage and workflow context
Dependency/runtime risk
WARN46
command execution surface, credential or environment access
Install availability
PASS92
npx skills add datadog-labs/agent-skills --skill agent-observability-eval-pipeline
Install command safety
INFO68
dynamic command execution, standard package or runtime install path
Permission surface
FAIL22
secrets or environment access, shell or command execution
Repository evidence
PASS86
https://github.com/datadog-labs/agent-skills/tree/main/agent-observability/agent-observability-eval-pipeline
Review status
INFO66
AI review data available
Agent Proven outcomes
INFO54
No agent outcome data yet
Checks
Install path
92
npx skills add datadog-labs/agent-skills --skill agent-observability-eval-pipeline
Repository
88
https://github.com/datadog-labs/agent-skills/tree/main/agent-observability/agent-observability-eval-pipeline
License
86
MIT
Maintenance
100
9d since push
AI review
55
The skill invokes sub-skills (agent-observability-session-classify, etc.) that are not included in this repository; they must be separately installed for the pipeline to work.
Warnings
Method
This report combines public metadata, AI review output, repository freshness, install readiness, OpenAgentSkill events, quality scoring, trust checks, and the agent safety gate. It is not a full source-code security review.
Compare nearby options
Guidance for distinctive, intentional UI design, typography, visual direction, and non-template-like product interfaces.
174K Stars · Audit report
Design and implementation guidance for distinctive landing pages, portfolios, product demos, and purposeful redesigns.
84K Stars · Audit report
Create original visual art, posters, PNG assets, and PDF documents through a clear design philosophy.
174K Stars · Audit report
OpenAgentSkill Trust Score
The Trust Score helps an agent decide whether a skill is safe enough to shortlist before installation.
GitHub adoption
INFO62
158 GitHub stars
Stars/forks activity
WARN57
158 stars, 25 forks; issue activity unavailable in current metadata
Recent maintenance
PASS100
9d since push
License clarity
PASS86
MIT
README/SKILL.md completeness
PASS86
Metadata includes enough usage and workflow context
Dependency/runtime risk
WARN46
command execution surface, credential or environment access
Install availability
PASS92
npx skills add datadog-labs/agent-skills --skill agent-observability-eval-pipeline
Install command safety
INFO68
dynamic command execution, standard package or runtime install path
Permission surface
FAIL22
secrets or environment access, shell or command execution
Repository evidence
PASS86
https://github.com/datadog-labs/agent-skills/tree/main/agent-observability/agent-observability-eval-pipeline
Review status
INFO66
AI review data available
Agent Proven outcomes
INFO54
No agent outcome data yet
Checks
Install path
92
npx skills add datadog-labs/agent-skills --skill agent-observability-eval-pipeline
Repository
88
https://github.com/datadog-labs/agent-skills/tree/main/agent-observability/agent-observability-eval-pipeline
License
86
MIT
Maintenance
100
9d since push
AI review
55
The skill invokes sub-skills (agent-observability-session-classify, etc.) that are not included in this repository; they must be separately installed for the pipeline to work.
Warnings
Method
This report combines public metadata, AI review output, repository freshness, install readiness, OpenAgentSkill events, quality scoring, trust checks, and the agent safety gate. It is not a full source-code security review.
Compare nearby options
Guidance for distinctive, intentional UI design, typography, visual direction, and non-template-like product interfaces.
174K Stars · Audit report
Design and implementation guidance for distinctive landing pages, portfolios, product demos, and purposeful redesigns.
84K Stars · Audit report
Create original visual art, posters, PNG assets, and PDF documents through a clear design philosophy.
174K Stars · Audit report
OpenAgentSkill Trust Score
The Trust Score helps an agent decide whether a skill is safe enough to shortlist before installation.
GitHub adoption
INFO62
158 GitHub stars
Stars/forks activity
WARN57
158 stars, 25 forks; issue activity unavailable in current metadata
Recent maintenance
PASS100
9d since push
License clarity
PASS86
MIT
README/SKILL.md completeness
PASS86
Metadata includes enough usage and workflow context
Dependency/runtime risk
WARN46
command execution surface, credential or environment access
Install availability
PASS92
npx skills add datadog-labs/agent-skills --skill agent-observability-eval-pipeline
Install command safety
INFO68
dynamic command execution, standard package or runtime install path
Permission surface
FAIL22
secrets or environment access, shell or command execution
Repository evidence
PASS86
https://github.com/datadog-labs/agent-skills/tree/main/agent-observability/agent-observability-eval-pipeline
Review status
INFO66
AI review data available
Agent Proven outcomes
INFO54
No agent outcome data yet
Checks
Install path
92
npx skills add datadog-labs/agent-skills --skill agent-observability-eval-pipeline
Repository
88
https://github.com/datadog-labs/agent-skills/tree/main/agent-observability/agent-observability-eval-pipeline
License
86
MIT
Maintenance
100
9d since push
AI review
55
The skill invokes sub-skills (agent-observability-session-classify, etc.) that are not included in this repository; they must be separately installed for the pipeline to work.
Warnings
Method
This report combines public metadata, AI review output, repository freshness, install readiness, OpenAgentSkill events, quality scoring, trust checks, and the agent safety gate. It is not a full source-code security review.
Compare nearby options
Guidance for distinctive, intentional UI design, typography, visual direction, and non-template-like product interfaces.
174K Stars · Audit report
Design and implementation guidance for distinctive landing pages, portfolios, product demos, and purposeful redesigns.
84K Stars · Audit report
Create original visual art, posters, PNG assets, and PDF documents through a clear design philosophy.
174K Stars · Audit report
OpenAgentSkill Trust Score
The Trust Score helps an agent decide whether a skill is safe enough to shortlist before installation.
GitHub adoption
INFO62
158 GitHub stars
Stars/forks activity
WARN57
158 stars, 25 forks; issue activity unavailable in current metadata
Recent maintenance
PASS100
9d since push
License clarity
PASS86
MIT
README/SKILL.md completeness
PASS86
Metadata includes enough usage and workflow context
Dependency/runtime risk
WARN46
command execution surface, credential or environment access
Install availability
PASS92
npx skills add datadog-labs/agent-skills --skill agent-observability-eval-pipeline
Install command safety
INFO68
dynamic command execution, standard package or runtime install path
Permission surface
FAIL22
secrets or environment access, shell or command execution
Repository evidence
PASS86
https://github.com/datadog-labs/agent-skills/tree/main/agent-observability/agent-observability-eval-pipeline
Review status
INFO66
AI review data available
Agent Proven outcomes
INFO54
No agent outcome data yet
Checks
Install path
92
npx skills add datadog-labs/agent-skills --skill agent-observability-eval-pipeline
Repository
88
https://github.com/datadog-labs/agent-skills/tree/main/agent-observability/agent-observability-eval-pipeline
License
86
MIT
Maintenance
100
9d since push
AI review
55
The skill invokes sub-skills (agent-observability-session-classify, etc.) that are not included in this repository; they must be separately installed for the pipeline to work.
Warnings
Method
This report combines public metadata, AI review output, repository freshness, install readiness, OpenAgentSkill events, quality scoring, trust checks, and the agent safety gate. It is not a full source-code security review.
Compare nearby options
Guidance for distinctive, intentional UI design, typography, visual direction, and non-template-like product interfaces.
174K Stars · Audit report
Design and implementation guidance for distinctive landing pages, portfolios, product demos, and purposeful redesigns.
84K Stars · Audit report
Create original visual art, posters, PNG assets, and PDF documents through a clear design philosophy.
174K Stars · Audit report
86
Usable description available
Dependency risk
46
command execution surface, credential or environment access
Install command safety
68
dynamic command execution, standard package or runtime install path
Permission surface
22
secrets or environment access, shell or command execution
Stars/forks activity
57
158 stars, 25 forks; issue activity unavailable in current metadata
Adoption
68
158 GitHub stars
Agent safety v2
This skill should not be selected by an agent without explicit human security review.
Do not auto-install. Inspect the source, dependencies, and permission surface first.
Shell or command execution
highSkill metadata references terminal, CLI, shell, subprocess, or command execution workflows.
Network access
mediumSkill likely fetches remote pages, APIs, repositories, or external services.
Filesystem access
mediumSkill may read or write project files, documents, generated artifacts, or local workspace state.
Secrets or environment access
highSkill metadata references credentials, tokens, environment variables, or secret-bearing workflows.
86
Usable description available
Dependency risk
46
command execution surface, credential or environment access
Install command safety
68
dynamic command execution, standard package or runtime install path
Permission surface
22
secrets or environment access, shell or command execution
Stars/forks activity
57
158 stars, 25 forks; issue activity unavailable in current metadata
Adoption
68
158 GitHub stars
Agent safety v2
This skill should not be selected by an agent without explicit human security review.
Do not auto-install. Inspect the source, dependencies, and permission surface first.
Shell or command execution
highSkill metadata references terminal, CLI, shell, subprocess, or command execution workflows.
Network access
mediumSkill likely fetches remote pages, APIs, repositories, or external services.
Filesystem access
mediumSkill may read or write project files, documents, generated artifacts, or local workspace state.
Secrets or environment access
highSkill metadata references credentials, tokens, environment variables, or secret-bearing workflows.
86
Usable description available
Dependency risk
46
command execution surface, credential or environment access
Install command safety
68
dynamic command execution, standard package or runtime install path
Permission surface
22
secrets or environment access, shell or command execution
Stars/forks activity
57
158 stars, 25 forks; issue activity unavailable in current metadata
Adoption
68
158 GitHub stars
Agent safety v2
This skill should not be selected by an agent without explicit human security review.
Do not auto-install. Inspect the source, dependencies, and permission surface first.
Shell or command execution
highSkill metadata references terminal, CLI, shell, subprocess, or command execution workflows.
Network access
mediumSkill likely fetches remote pages, APIs, repositories, or external services.
Filesystem access
mediumSkill may read or write project files, documents, generated artifacts, or local workspace state.
Secrets or environment access
highSkill metadata references credentials, tokens, environment variables, or secret-bearing workflows.
86
Usable description available
Dependency risk
46
command execution surface, credential or environment access
Install command safety
68
dynamic command execution, standard package or runtime install path
Permission surface
22
secrets or environment access, shell or command execution
Stars/forks activity
57
158 stars, 25 forks; issue activity unavailable in current metadata
Adoption
68
158 GitHub stars
Agent safety v2
This skill should not be selected by an agent without explicit human security review.
Do not auto-install. Inspect the source, dependencies, and permission surface first.
Shell or command execution
highSkill metadata references terminal, CLI, shell, subprocess, or command execution workflows.
Network access
mediumSkill likely fetches remote pages, APIs, repositories, or external services.
Filesystem access
mediumSkill may read or write project files, documents, generated artifacts, or local workspace state.
Secrets or environment access
highSkill metadata references credentials, tokens, environment variables, or secret-bearing workflows.