OPENAGENTSKILL / DIRECTORY

AI Agent Skills

다음 작업에 맞는 스킬을 찾아보세요. Codex, Claude Code, Cursor 등을 지원합니다.

검색 결과 · “llm-eval”

110 Skills

검색 결과: 110

AI 및 지식

No visual example yet

Explore the skill

非线智能 NoneLinear - ReLE评测:中文AI大模型能力评测(持续更新):目前已囊括374个大模型,覆盖chatgpt、gpt-5.4、谷歌gemini-3.1-pro、Claude-4.6、文心ERNIE-X1.1、ERNIE-5.0、qwen3.6-max、qwen3.6-plus、百川、讯飞星火、商汤senseChat…

가격 미확인AI 및 지식Claude CodeOpenAI Agents사용 전 검토
6.2천GitHub
AI 및 지식

No visual example yet

Explore the skill

Langfuse

langfuse

🪢 Open source AI engineering platform: LLM evals, observability, metrics, prompt management, playground, datasets. Integrates with OpenTelemetry, LangChain, OpenAI SDK,…

가격 미확인AI 및 지식OpenAI AgentsLangChain사용 전 검토
2.9만GitHub
개발 및 테스트

No visual example yet

Explore the skill

Opik

comet-ml

Debug, evaluate, and monitor your LLM applications, RAG systems, and agentic workflows with comprehensive tracing, automated evaluations, and production-ready dashboards.

가격 미확인개발 및 테스트LangChain사용 전 검토
2만GitHub
AI 및 지식

No visual example yet

Explore the skill

Agenta

Agenta-AI

The open-source LLMOps platform: prompt playground, prompt management, LLM evaluation, and LLM observability all in one place.

가격 미확인AI 및 지식사용 전 검토
4.5천GitHub

가이드 및 비교

개발자용