OpenAgentSkill Registry Manifest Skill: promptfoo-evals Slug: promptfoo-promptfoo-evals Category: design-creative Description: Write, refine, run, and QA non-redteam promptfoo eval suites after the target or provider already works: prompts, vars, test cases, assertions, model-graded rubrics, transforms, datasets, output exports, filters, and CI gates. Use for regression tests and eval-suite authoring. Do not use for connecting a new target/provider, mapping HTTP requests or auth, smoke-testing an endpoint, or redteam plugin/strategy setup; use `promptfoo-provider-setup` for connection work instead. Agent fit: - Decision: 100/100 Production-ready - Primary fit: Research agents - Role: Primary pick Supply profile: - Track: Research and knowledge work - Scenario: Research agents - Applicable agents: Claude Code, OpenAI Agents, CLI, Codex, Cursor - Maintenance: 3d since push - Risk: Needs review Trust: - Trust score: 78/100 Strong shortlist - Audit: 86/100 Needs review Attribution: - Status: Registry indexed - Source: github fast track - Creator: promptfoo - Claim URL: https://www.openagentskill.com/skills/promptfoo-promptfoo-evals#claim-this-skill Install: npx skills add promptfoo/promptfoo --skill promptfoo-evals URLs: - Web: https://www.openagentskill.com/skills/promptfoo-promptfoo-evals - API: https://www.openagentskill.com/api/agent/skills/promptfoo-promptfoo-evals - Install API: https://www.openagentskill.com/api/skills/promptfoo-promptfoo-evals/install - Repository: https://github.com/promptfoo/promptfoo/tree/main/.claude/skills/promptfoo-evals