OPENAGENTSKILL / DIRECTORY

AI Agent Skills

Encuentra una habilidad para tu próxima tarea con Codex, Claude Code, Cursor y más.

Resultados · “evaluation”

12 Skills

Resultados: 12

IA y conocimiento

No visual example yet

Explore the skill

Sharing both practical insights and theoretical knowledge about LLM evaluation that we gathered while managing the Open LLM Leaderboard and designing lighteval!

Precio sin confirmarIA y conocimientoRevisar antes de usar
2,1 milGitHub
IA y conocimiento

No visual example yet

Explore the skill

Agenta

Agenta-AI

The open-source LLMOps platform: prompt playground, prompt management, LLM evaluation, and LLM observability all in one place.

Precio sin confirmarIA y conocimientoRevisar antes de usar
4,5 milGitHub
IA y conocimiento

No visual example yet

Explore the skill

Skills

langfuse

Agent Skills for Langfuse, the open source LLM engineering platform for tracing, prompt management, and evaluation

Precio sin confirmarIA y conocimientoClaude CodeRevisar antes de usar
214GitHub
IA y conocimiento

No visual example yet

Explore the skill

🦄 Unitxt is a Python library for enterprise-grade evaluation of AI performance, offering the world's largest catalog of tools and data for end-to-end AI benchmarking

Precio sin confirmarIA y conocimientoRevisar antes de usar
214GitHub
IA y conocimiento

No visual example yet

Explore the skill

Patterns and techniques for evaluating and improving AI agent outputs. Use this skill when: - Implementing self-critique and reflection loops - Building evaluator-optimi…

Precio sin confirmarIA y conocimientoClaude Code
38,5 milGitHub
IA y conocimiento

No visual example yet

Explore the skill

RAG Fusion

Raudaschl

RAG-Fusion: multi-query generation + Reciprocal Rank Fusion for better retrieval-augmented generation. Includes evaluation harness with NFCorpus/BEIR.

Precio sin confirmarIA y conocimientoOpenAI AgentsRevisar antes de usar
946GitHub

Guías y comparativas

Para desarrolladores