Cleverhans
cleverhans-lab
An adversarial example library for constructing attacks, building defenses, and benchmarking both
OPENAGENTSKILL / DIRECTORY
Find a skill for your next task. Explore tools for Codex, Claude Code, Cursor and more.
1–4 / 4
Candidates in this shortlist, not the full registry. GitHub stars belong to repositories, not individual skills.
Results: 4
cleverhans-lab
An adversarial example library for constructing attacks, building defenses, and benchmarking both
rentruewang
Bayesian Optimization as a Coverage Tool for Evaluating LLMs. Accurate evaluation (benchmarking) that's 10 times faster with just a few lines of modular code.
IBM
🦄 Unitxt is a Python library for enterprise-grade evaluation of AI performance, offering the world's largest catalog of tools and data for end-to-end AI benchmarking
bark-simulator
Open-Source Framework for Development, Simulation and Benchmarking of Behavior Planning Algorithms for Autonomous Driving