OPENAGENTSKILL / DIRECTORY

AI Agent Skills

Encuentra una habilidad para tu próxima tarea con Codex, Claude Code, Cursor y más.

Resultados · “benchmarking”

44 Skills

Resultados: 44

Diseño y UI

No visual example yet

Explore the skill

Design and validity review for studies that benchmark one or more AI systems against a human-expert panel as the reference. Covers the evaluation question and arm defini…

Precio sin confirmarDiseño y UIClaude Code
291GitHub
IA y conocimiento

No visual example yet

Explore the skill

Bocoel

rentruewang

Bayesian Optimization as a Coverage Tool for Evaluating LLMs. Accurate evaluation (benchmarking) that's 10 times faster with just a few lines of modular code.

Precio sin confirmarIA y conocimientoRevisar antes de usar
289GitHub
IA y conocimiento

No visual example yet

Explore the skill

Agentops

AgentOps-AI

Python SDK for AI agent monitoring, LLM cost tracking, benchmarking, and more. Integrates with most LLMs and agent frameworks including CrewAI, Agno, OpenAI Agents SDK,…

Precio sin confirmarIA y conocimientoClaude CodeOpenAI AgentsRevisar antes de usar
5,8 milGitHub
Vídeo y audio

No visual example yet

Explore the skill

Video ChatGPT

mbzuai-oryx

[ACL 2024 🔥] Video-ChatGPT is a video conversation model capable of generating meaningful conversation about videos. It combines the capabilities of LLMs with a pretrai…

Precio sin confirmarVídeo y audioOpenAI AgentsRevisar antes de usar
1,5 milGitHub
Desarrollo

No visual example yet

Explore the skill

Golang benchmarking, profiling, and performance measurement. Use when writing, running, or comparing Go benchmarks, profiling hot paths with pprof, interpreting CPU/memo…

Precio sin confirmarDesarrolloClaude Code
3 milGitHub

Guías y comparativas

Para desarrolladores