Direktori skill

Temukan skill yang dapat digunakan kembali untuk AI agents.

Cari skill GitHub nyata berdasarkan tugas lalu periksa stars, trust, audit, kategori, dan jalur pemasangan sebelum digunakan.

Setiap rekomendasi tetap terhubung dengan repositori, audit, dan jalur pemasangannya.

Hasil pencarian: benchmark-report

Direktori bahasa Inggris

A Python library for anomaly detection across tabular, time series, graph, text, and image data. 60+ detectors, benchmark-backed ADEngine orchestration, and an agentic workflow for AI agents.

9.9K
Stars
86/100
Kepercayaan
Kategori: ml-automationAudit

Checks whether Kubernetes is deployed according to security best practices as defined in the CIS Kubernetes Benchmark

8.1K
Stars
86/100
Kepercayaan
Kategori: devopsAudit

✨ The agentic HTML editor — your local AI agent writes the HTML, you ship it. 🚀 75 Skills × 9 Surfaces (magazine · deck · poster · XHS / tweet · prototype · data report · Hyperframes) 🛡️ Sandboxed preview · 📤 1-click to WeChat / X / Zhihu / HTML / PNG 🔑 Zero API key — Claude Code / Cursor / Codex / Gemini / Copilot / OpenCode / Qwen / Aider.

6.7K
Stars
76/100
Kepercayaan
Kategori: agent-skillsAudit

非线智能 NoneLinear - ReLE评测:中文AI大模型能力评测(持续更新):目前已囊括374个大模型,覆盖chatgpt、gpt-5.4、谷歌gemini-3.1-pro、Claude-4.6、文心ERNIE-X1.1、ERNIE-5.0、qwen3.6-max、qwen3.6-plus、百川、讯飞星火、商汤senseChat等商用模型, 以及step3.5-flash、kimi-k2.6、ernie4.5、MiniMax-M2.7、deepseek-v4、Qwen3.6、llama4、智谱GLM-5.1、MiMo-V2、LongCat、gemma4、mistral等开源大模型。不仅提供排行榜,也提供规模超200万的大模型缺陷库!方便广大社区研究分析、改进大模型。

6.2K
Stars
77/100
Kepercayaan
Kategori: agent-frameworksAudit

Vendor-agnostic orchestration for training, inference and agentic workloads across NVIDIA, AMD, TPU, and Tenstorrent on clouds, Kubernetes, and bare metal.

2.2K
Stars
85/100
Kepercayaan
Kategori: agent-skillsAudit

A Codex skill that analyzes startup URLs or product ideas to find evidence-backed potential first customers using public signals.

989
Stars
84/100
Kepercayaan
Kategori: marketing-growthAudit

35 production-grade agentic AI architectures (Reflexion, LATS, GraphRAG, MemGPT, Voyager, BrowserAgent, ...) — a Python library and runnable textbook with multi-provider LLM support and a 17-task benchmark leaderboard.

3.7K
Stars
85/100
Kepercayaan
Kategori: agent-frameworksAudit

AI-powered bug bounty hunting from your terminal - recon, 20 vuln classes, autonomous hunting, and report generation. All inside Claude Code.

3.5K
Stars
79/100
Kepercayaan
Kategori: securityAudit

MTEB: Massive Text Embedding Benchmark

3.3K
Stars
80/100
Kepercayaan
Kategori: rag-knowledgeAudit

Help your coding agents (Claude Code, Codex, Qoder, Cursor, and other coding agents) get better at getting better.

1.1K
Stars
86/100
Kepercayaan
Kategori: utilityAudit

A skill for AI agents (Claude Code, Codex, Cursor) that rewrites Traditional Chinese text to remove AI writing patterns, correct China-Taiwan localization, and fix punctuation.

691
Stars
83/100
Kepercayaan
Kategori: utilityAudit

A Claude Code skill bundle for bug hunting and external red-team work — 71 skills, 15 slash commands, 681 disclosed-report patterns curated across 24 core vulnerability classes, plus enterprise identity + infrastructure attack matrices.

2.2K
Stars
76/100
Kepercayaan
Kategori: developmentAudit