Annuaire de skills

Découvrez des skills réutilisables pour les AI agents.

Recherchez de vrais skills GitHub par tâche et vérifiez Stars, confiance, audit, catégorie et chemin d’installation avant de les utiliser.

Chaque recommandation reste clairement reliée à son dépôt, son audit et son chemin d’installation.

Résultats de recherche: evaluation-metrics

Annuaire en anglais

The open and composable observability and data visualization platform. Visualize metrics, logs, and traces from multiple sources like Prometheus, Loki, Elasticsearch, InfluxDB, Postgres and many more.

74K
Stars
86/100
Confiance
Catégorie: data-analysisAudit

A comprehensive set of 38 marketing skills and 5 commands for Claude Code covering SEO/GEO and influencer marketing with evaluation frameworks.

2.6K
Stars
86/100
Confiance
Catégorie: productivityAudit

Automated auditing, performance metrics, and best practices for the web.

30K
Stars
82/100
Confiance
Catégorie: coding-agentsAudit

🪢 Open source AI engineering platform: LLM evals, observability, metrics, prompt management, playground, datasets. Integrates with OpenTelemetry, LangChain, OpenAI SDK, LiteLLM, and more. 🍊YC W23

29K
Stars
79/100
Confiance
Catégorie: developmentAudit

SigNoz is an open-source observability platform native to OpenTelemetry with logs, traces and metrics in a single application. An open-source alternative to DataDog, NewRelic, etc. 🔥 🖥. 👉 Open source Application Performance Monitoring (APM) & Observability tool

27K
Stars
79/100
Confiance
Catégorie: devopsAudit

A Codex skill for generating minimal zine-style editorial poster prompts and images.

6.3K
Stars
83/100
Confiance
Catégorie: design-creativeAudit

Open source observability platform for logs, metrics, traces, frontend monitoring, pipelines and LLM observability. A sophisticated, simple and highly performant alternative to Datadog, Splunk, and Elasticsearch with 140x lower storage costs and single binary deployment.

19K
Stars
87/100
Confiance
Catégorie: devopsAudit

📊 An infographics generator with 30+ plugins and 300+ options to display stats about your GitHub account and render them as SVG, Markdown, PDF or JSON!

17K
Stars
80/100
Confiance
Catégorie: automationAudit

Open-source evaluation toolkit of large multi-modality models (LMMs), support 220+ LMMs, 80+ benchmarks

4.2K
Stars
85/100
Confiance
Catégorie: robotics-iotAudit

BISHENG is an open LLM devops platform for next generation Enterprise AI applications. Powerful and comprehensive features include: GenAI workflow, RAG, Agent, Unified model management, Evaluation, SFT, Dataset Management, Enterprise-level System Management, Observability and more.

11K
Stars
87/100
Confiance
Catégorie: document-processingAudit

AI Observability & Evaluation

10K
Stars
76/100
Confiance
Catégorie: developmentAudit

The open-source LLMOps platform: prompt playground, prompt management, LLM evaluation, and LLM observability all in one place.

4.5K
Stars
78/100
Confiance
Catégorie: developmentAudit