OPENAGENTSKILL / DIRECTORY

AI Agent Skills

다음 작업에 맞는 스킬을 찾아보세요. Codex, Claude Code, Cursor 등을 지원합니다.

검색 결과 · “benchmarking”

44 Skills

검색 결과: 44

디자인 및 UI

No visual example yet

Explore the skill

Design and validity review for studies that benchmark one or more AI systems against a human-expert panel as the reference. Covers the evaluation question and arm defini…

가격 미확인디자인 및 UIClaude Code
291GitHub
AI 및 지식

No visual example yet

Explore the skill

Bocoel

rentruewang

Bayesian Optimization as a Coverage Tool for Evaluating LLMs. Accurate evaluation (benchmarking) that's 10 times faster with just a few lines of modular code.

가격 미확인AI 및 지식사용 전 검토
289GitHub
AI 및 지식

No visual example yet

Explore the skill

Agentops

AgentOps-AI

Python SDK for AI agent monitoring, LLM cost tracking, benchmarking, and more. Integrates with most LLMs and agent frameworks including CrewAI, Agno, OpenAI Agents SDK,…

가격 미확인AI 및 지식Claude CodeOpenAI Agents사용 전 검토
5.8천GitHub
영상 및 오디오

No visual example yet

Explore the skill

Video ChatGPT

mbzuai-oryx

[ACL 2024 🔥] Video-ChatGPT is a video conversation model capable of generating meaningful conversation about videos. It combines the capabilities of LLMs with a pretrai…

가격 미확인영상 및 오디오OpenAI Agents사용 전 검토
1.5천GitHub
개발 및 테스트

No visual example yet

Explore the skill

Golang benchmarking, profiling, and performance measurement. Use when writing, running, or comparing Go benchmarks, profiling hot paths with pprof, interpreting CPU/memo…

가격 미확인개발 및 테스트Claude Code
3천GitHub

가이드 및 비교

개발자용