No visual example yet
Explore the skillKube Bench
aquasecurity
Checks whether Kubernetes is deployed according to security best practices as defined in the CIS Kubernetes Benchmark
OPENAGENTSKILL / DIRECTORY
Find a skill for your next task. Explore tools for Codex, Claude Code, Cursor and more.
25 Skills
Results: 25
No visual example yet
Explore the skillaquasecurity
Checks whether Kubernetes is deployed according to security best practices as defined in the CIS Kubernetes Benchmark
No visual example yet
Explore the skillMLSysOps
🤖 MLE-Agent: Your intelligent companion for seamless AI engineering and research. 🔍 Integrate with arxiv and paper with code to provide better code/research plans 🧰 O…
No visual example yet
Explore the skillK-Dense-AI
Autonomously improve a real artifact (code, training recipe, agent harness, data pipeline, prompt) against an objective and an evaluator, using Hypothesis Tree Refinemen…
No visual example yet
Explore the skillAyanami0730
DeepResearch Bench: A Comprehensive Benchmark for Deep Research Agents
No visual example yet
Explore the skillonyx-dot-app
Dataset and benchmark for RAG on company internal documents.
No visual example yet
Explore the skillalibaba
An Alibaba open-source multi-language benchmark for evaluating LLMs in repository-level automatic code review, featuring an AI-assisted and expert-verified dataset.
No visual example yet
Explore the skillAndrejOrsula
Robot Learning Beyond Earth
No visual example yet
Explore the skillb7leung
200+ detailed flashcards useful for reviewing topics in machine learning, computer vision, and computer science.
No visual example yet
Explore the skillarthur-ai
No visual example yet
Explore the skillVisko-Platform
VEFX-Bench: A Holistic Benchmark for Generic Video Editing and Visual Effects
No visual example yet
Explore the skillagentii-ai
Med peer benchmarking: select actual biotech/pharma peers via the med universe (drug/indication overlap where possible) and compare med-relevant metrics — pipeline depth…
No visual example yet
Explore the skillagentii-ai
Med peer benchmarking: select actual biotech/pharma peers via the med universe (drug/indication overlap where possible) and compare med-relevant metrics — pipeline depth…
No visual example yet
Explore the skillintel
Benchmark a **running SGLang-XPU server** on an Intel GPU using `sglang.bench_serving`. Measures TTFT, TPOT, ITL, end-to-end latency, and throughput against the OpenAI-c…
No visual example yet
Explore the skillKodezi
Kodezi Chronos is a debugging-first language model that achieves state-of-the-art results on SWE-bench Lite (80.33%) and 67% real-world fix accuracy, over six times bett…
No visual example yet
Explore the skillJARVIS-Xs
SE-Agent is a self-evolution framework for LLM Code agents. It enables trajectory-level evolution to exchange information across reasoning paths via Revision, Recombinat…
No visual example yet
Explore the skillSWE-agent
The 100 line AI agent that solves GitHub issues or helps you in your command line. Radically simple, no huge configs, no giant monorepo—but scores >74% on SWE-bench veri…