Vorinstallations-Eval
Llm App Eval-Bericht.
Eine maschinenlesbare Installationsentscheidung für Agents: Task-Fit, Trust Score, Audit Score, Installationssicherheit, Berechtigungsumfang und ein konkreter Validierungsplan vor dem Einsatz im Workspace.
Manuelle Prüfung
Allow agent install in a sandbox or low-risk workspace, then promote after one successful narrow task.
Erforderliche Gates
Prüfungen, die ein Agent vor der Installation bestehen muss
Aufgabenpassung
70
Task fit is weak; compare alternatives before selecting.
- Evaluate Llm App before installing it in an agent workflow
- Daten
- RAG and knowledge workflows; Claude Code teams; teams that value GitHub adoption signals
Installationspfad
92
Install handoff is available.
- npx skills add pathwaycom/llm-app
Sicherheit des Installationsbefehls
92
Standard-Paket- oder Laufzeit-Installationspfad
- npx skills add pathwaycom/llm-app
Vertrauenswert
92
Strong OpenAgentSkill Trust Score across adoption, recent maintenance, license clarity, documentation, dependency/runtime risk, install safety, permission surface, and install availability.
- Produktionskandidat
- 59K GitHub-Stars
- MIT
Audit-Score
93
Sicher zu testen
- No major audit warning from metadata.
Agent-Sicherheitsprüfung
89
Strong metadata, audit, install, and review signals. Suitable for agent shortlists after normal workspace review.
- Allow agent install in a sandbox or low-risk workspace, then promote after one successful narrow task.
- Verified listing
Lizenzklarheit
86
MIT
- MIT
Berechtigungsumfang
100
Keine Hochrisiko-Berechtigungsfläche in öffentlichen Metadaten
- Network access: medium
Validierungsplan
Was der Agent als Nächstes tun sollte
- 1Inspect repository, README/SKILL.md, license, and recent commits before production use.
- 2Install in an isolated workspace or sandbox with no production secrets available.
- 3Run the smallest representative task and record files touched, commands run, network access, and outputs.
- 4Compare the selected skill against at least one alternative when the eval status is review or failed.
- 5Promote only after the agent reports a successful verification result and unresolved warnings are accepted.
Nicht verwenden, wenn
Bedingungen für eine andere Skill
- Teams, die ein vom Anbieter unterstütztes SLA benötigen
- Hochregulierte Umgebungen ohne interne Sicherheitsprüfung
- No major risk signals from current metadata
- No major trust warnings detected from available metadata
- Production credentials, payments, or irreversible account changes without explicit human review
- Sensitive private data before reviewing repository code, license, and permission surface
Unterstützende Prüfungen
Vertrauenssignale hinter der Entscheidung
README/SKILL.md-Vollständigkeit
Bestanden90
Metadaten enthalten ausreichend Nutzungs- und Workflow-Kontext
Aktuelle Wartung
Bestanden88
2 Monate seit dem letzten Push
Alternativen verfügbar
Bestanden82
Alternative skills are available for comparison.