agent-evaluation-and-guardrails
selvarajmurugesan90
Guides building evaluation harnesses, regression test suites, and runtime guardrails for LLM agents. Use when a user asks to "evaluate this agent," "write test cases for…
OPENAGENTSKILL / DIRECTORY
Find a skill for your next task. Explore tools for Codex, Claude Code, Cursor and more.
Find a skill for your next task. Preview examples where available.
Page 31 · 5 shown · 485 public entries
Results: 485
selvarajmurugesan90
Guides building evaluation harnesses, regression test suites, and runtime guardrails for LLM agents. Use when a user asks to "evaluate this agent," "write test cases for…
mralaminahamed
Use when setting up PHPCS with WordPress Coding Standards (WPCS), configuring phpcs.xml.dist, running phpcs/phpcbf, fixing sniff violations, adding PHPCS to CI (GitHub A…
ZealynxSecurity
Deep analysis of a single security check against local Solidity code. Use with framework check IDs like AC-01, LN-02, VT-03. Also works during manual assessment to get A…
ZealynxSecurity
Quick security scan of Solidity smart contracts. Analyzes code against real audit findings from 4,500+ Solodit references across 39 DeFi verticals. Use when reviewing So…
ishandutta2007
Subjects every non-trivial decision to a fresh-context adversarial review before it stands. Use when correctness matters more than speed, when working in unfamiliar code…