story-html-publisher Eval
=========================
Status: failed
Score: 67/100
Risk: high
Decision: do_not_auto_install
Policy: block
Reason: Audit score: Risky
Install:
npx skills add hassancs91/claude-image-generation --skill story-html-publisher
Required checks:
- PASS Task fit: Task wording matches this skill metadata.
- PASS Install path: Install handoff is available.
- PASS Install command safety: standard package or runtime install path
- WARN Trust score: Potentially useful, but at least one trust signal needs human inspection.
- FAIL Audit score: Risky
- FAIL Agent safety gate: This skill should not be selected by an agent without explicit human security review.
- PASS License clarity: MIT
- FAIL Permission surface: shell or command execution, filesystem or document access
Warnings:
- Trust score: Potentially useful, but at least one trust signal needs human inspection.
- Audit risk risky exceeds max_risk=medium
- High-risk permission hints: Shell or command execution
- Permission surface may require sandboxing
- Financial research output is not financial advice; require human review before any live investment decision
- Potential broker, wallet, exchange, or real-money execution surface; sandbox and explicit approval are required
- The build script downloads media from URLs without explicit validation, which could be a vector for SSRF if upstream artifacts are untrusted. However, the skill is designed for a controlled pipeline.
- The skill does not include its own license file, but the repository is MIT licensed, which is acceptable.
- Financial research output is not financial advice; require human review before any live investment decision.
- This skill may touch real-money trading, broker, wallet, or exchange operations; use only in a sandbox with explicit approval.
- Quality score needs review
- Permission surface needs review: shell or command execution, filesystem or document access
Validation plan:
1. Inspect repository, README/SKILL.md, license, and recent commits before production use.
2. Install in an isolated workspace or sandbox with no production secrets available.
3. Run the smallest representative task and record files touched, commands run, network access, and outputs.
4. Compare the selected skill against at least one alternative when the eval status is review or failed.
5. Promote only after the agent reports a successful verification result and unresolved warnings are accepted.
Do not use when:
- teams that need a vendor-supported SLA
- production agents without a repository review
- The build script downloads media from URLs without explicit validation, which could be a vector for SSRF if upstream artifacts are untrusted. However, the skill is designed for a controlled pipeline.
- No OpenAgentSkill engagement data yet
- Audit risk risky exceeds max_risk=medium
- High-risk permission hints: Shell or command execution
- Permission surface may require sandboxing
- Financial research output is not financial advice; require human review before any live investment decision
URLs:
- Skill: https://www.openagentskill.com/skills/hassancs91-story-html-publisher
- Audit: https://www.openagentskill.com/skills/hassancs91-story-html-publisher/audit
- JSON: https://www.openagentskill.com/api/agent/evals?slug=hassancs91-story-html-publisher