Direktori skill

Temukan skill yang dapat digunakan kembali untuk AI agents.

Cari skill GitHub nyata berdasarkan tugas lalu periksa stars, trust, audit, kategori, dan jalur pemasangan sebelum digunakan.

Setiap rekomendasi tetap terhubung dengan repositori, audit, dan jalur pemasangannya.

Hasil pencarian: evaluation

Direktori bahasa Inggris

A comprehensive set of 38 marketing skills and 5 commands for Claude Code covering SEO/GEO and influencer marketing with evaluation frameworks.

2.6K
Stars
86/100
Kepercayaan
Kategori: productivityAudit

A Codex skill for generating minimal zine-style editorial poster prompts and images.

6.3K
Stars
83/100
Kepercayaan
Kategori: design-creativeAudit

Open-source evaluation toolkit of large multi-modality models (LMMs), support 220+ LMMs, 80+ benchmarks

4.2K
Stars
85/100
Kepercayaan
Kategori: robotics-iotAudit

BISHENG is an open LLM devops platform for next generation Enterprise AI applications. Powerful and comprehensive features include: GenAI workflow, RAG, Agent, Unified model management, Evaluation, SFT, Dataset Management, Enterprise-level System Management, Observability and more.

11K
Stars
87/100
Kepercayaan
Kategori: document-processingAudit

AI Observability & Evaluation

10K
Stars
76/100
Kepercayaan
Kategori: developmentAudit

The open-source LLMOps platform: prompt playground, prompt management, LLM evaluation, and LLM observability all in one place.

4.5K
Stars
78/100
Kepercayaan
Kategori: developmentAudit

28 eval-informed mental models and critical-thinking skills for Claude Code, GitHub Copilot, Codex, Cursor, and other Agent Skills-compatible tools

941
Stars
84/100
Kepercayaan
Kategori: utilityAudit

Ship AI Agents to Google Cloud in minutes, not months. Production-ready templates with built-in CI/CD, evaluation, and observability.

6.5K
Stars
86/100
Kepercayaan
Kategori: developmentAudit

๐Ÿข Open-Source Evaluation & Testing library for LLM Agents

5.7K
Stars
81/100
Kepercayaan
Kategori: developmentAudit

Next-generation AI Agent Optimization Platform: Cozeloop addresses challenges in AI agent development by providing full-lifecycle management capabilities from development, debugging, and evaluation to monitoring.

5.6K
Stars
86/100
Kepercayaan
Kategori: developmentAudit

AutoRAG: An Open-Source Framework for Retrieval-Augmented Generation (RAG) Evaluation & Optimization with AutoML-Style Automation

4.8K
Stars
84/100
Kepercayaan
Kategori: dataAudit

Evaluation and Tracking for LLM Experiments and AI Agents

3.4K
Stars
80/100
Kepercayaan
Kategori: developmentAudit