OPENAGENTSKILL / DIRECTORY

AI Agent Skills

다음 작업에 맞는 스킬을 찾아보세요. Codex, Claude Code, Cursor 등을 지원합니다.

검색 결과 · “evaluation”

50 Skills

검색 결과: 50

AI 및 지식

No visual example yet

Explore the skill

Sharing both practical insights and theoretical knowledge about LLM evaluation that we gathered while managing the Open LLM Leaderboard and designing lighteval!

가격 미확인AI 및 지식사용 전 검토
2.1천GitHub
하드웨어 및 IoT

No visual example yet

Explore the skill

VLMEvalKit

open-compass

Open-source evaluation toolkit of large multi-modality models (LMMs), support 220+ LMMs, 80+ benchmarks

가격 미확인하드웨어 및 IoTClaude CodeOpenAI Agents사용 전 검토
4.2천GitHub
AI 및 지식

No visual example yet

Explore the skill

Agenta

Agenta-AI

The open-source LLMOps platform: prompt playground, prompt management, LLM evaluation, and LLM observability all in one place.

가격 미확인AI 및 지식사용 전 검토
4.5천GitHub
개발 및 테스트

No visual example yet

Explore the skill

Coze Loop

coze-dev

Next-generation AI Agent Optimization Platform: Cozeloop addresses challenges in AI agent development by providing full-lifecycle management capabilities from developmen…

가격 미확인개발 및 테스트사용 전 검토
5.6천GitHub
브라우저 및 자동화

No visual example yet

Explore the skill

AB3DMOT

xinshuoweng

(IROS 2020, ECCVW 2020) Official Python Implementation for "3D Multi-Object Tracking: A Baseline and New Evaluation Metrics"

가격 미확인브라우저 및 자동화사용 전 검토
1.8천GitHub
AI 및 지식

No visual example yet

Explore the skill

Skills

langfuse

Agent Skills for Langfuse, the open source LLM engineering platform for tracing, prompt management, and evaluation

가격 미확인AI 및 지식Claude Code사용 전 검토
214GitHub
AI 및 지식

No visual example yet

Explore the skill

🦄 Unitxt is a Python library for enterprise-grade evaluation of AI performance, offering the world's largest catalog of tools and data for end-to-end AI benchmarking

가격 미확인AI 및 지식사용 전 검토
214GitHub

가이드 및 비교

개발자용