Evaluasi sebelum pemasangan

gpt-image-2 Laporan evaluasi.

Keputusan pemasangan yang dapat dibaca mesin untuk Agent: kecocokan tugas, Trust Score, Audit Score, keamanan pemasangan, cakupan izin, dan rencana validasi konkret sebelum skill menyentuh workspace.

GagalRisiko tinggiBlokir Kebijakan
60
Laporan evaluasi
63
Kepercayaan
72
Audit
32
Keamanan Agent

do not auto install

Agent safety gate: This skill should not be selected by an agent without explicit human security review.

Gerbang wajib

Pemeriksaan yang harus dilalui Agent sebelum memasang

Buka JSON

Kecocokan tugas

84

Lulus

Task wording matches this skill metadata.

  • Evaluate gpt-image-2 before installing it in an agent workflow
  • creative
  • Design and creative workflows; Claude Code teams; builders willing to evaluate younger projects

Jalur pemasangan

92

Lulus

Install handoff is available.

  • npx skills add wubin1836/ai-hive-agent-skills --skill gpt-image-2

Keamanan perintah pemasangan

92

Lulus

Jalur pemasangan paket atau runtime standar

  • npx skills add wubin1836/ai-hive-agent-skills --skill gpt-image-2

Skor kepercayaan

63

Peringatan

Potentially useful, but at least one trust signal needs human inspection.

  • Tinjauan manual
  • 2 star GitHub
  • MIT

Skor audit

72

Peringatan

Perlu ditinjau

  • Dependency or permission surface needs review

Gerbang keamanan Agent

32

Gagal

This skill should not be selected by an agent without explicit human security review.

  • Do not auto-install. Inspect the source, dependencies, and permission surface first.
  • Metadata combines secrets access with shell or command execution

Kejelasan lisensi

86

Lulus

MIT

  • MIT

Cakupan izin

36

Gagal

secrets or environment access, shell or command execution

  • Shell or command execution: high
  • Network access: medium
  • Secrets or environment access: high

Rencana validasi

Langkah Agent berikutnya

  1. 1Inspect repository, README/SKILL.md, license, and recent commits before production use.
  2. 2Install in an isolated workspace or sandbox with no production secrets available.
  3. 3Run the smallest representative task and record files touched, commands run, network access, and outputs.
  4. 4Compare the selected skill against at least one alternative when the eval status is review or failed.
  5. 5Promote only after the agent reports a successful verification result and unresolved warnings are accepted.

Jangan gunakan ketika

Kondisi yang membutuhkan skill lain

  • Tim yang membutuhkan SLA dengan dukungan vendor
  • production agents without a repository review
  • Low GitHub adoption signal
  • SKILL.md is truncated in the excerpt, but the provided content is clear and well-structured; missing sections on error handling and limitations are not critical.
  • Petunjuk izin berisiko tinggi: Shell or command execution, Secrets or environment access
  • Dependency or permission surface needs review

Pemeriksaan pendukung

Sinyal kepercayaan di balik keputusan

Kelengkapan README/SKILL.md

Lulus

94

Metadata memuat konteks penggunaan dan alur kerja yang cukup

Pemeliharaan terbaru

Lulus

100

Diperbarui hari ini

Alternatif tersedia

Lulus

82

Alternative skills are available for comparison.