No visual example yet
Explore the skillagentic-eval
github
Patterns and techniques for evaluating and improving AI agent outputs. Use this skill when: - Implementing self-critique and reflection loops - Building evaluator-optimi…
OPENAGENTSKILL / DIRECTORY
Encuentra una habilidad para tu próxima tarea con Codex, Claude Code, Cursor y más.
3 Skills
Resultados: 3
No visual example yet
Explore the skillgithub
Patterns and techniques for evaluating and improving AI agent outputs. Use this skill when: - Implementing self-critique and reflection loops - Building evaluator-optimi…
No visual example yet
Explore the skillK-Dense-AI
Autonomously improve a real artifact (code, training recipe, agent harness, data pipeline, prompt) against an objective and an evaluator, using Hypothesis Tree Refinemen…
No visual example yet
Explore the skillTheGreenCedar
A codex plugin for running optimization loops inside a codebase. It is useful when you have a measurable target and many possible changes to try: test runtime, build spe…