OPENAGENTSKILL / DIRECTORY

AI Agent Skills

次のタスクに合うスキルを。Codex、Claude Code、Cursor などのツールを探せます。

検索結果 · “llm-evaluation”

109 Skills

検索結果: 109

AI・知識ベース

No visual example yet

Explore the skill

非线智能 NoneLinear - ReLE评测:中文AI大模型能力评测(持续更新):目前已囊括374个大模型,覆盖chatgpt、gpt-5.4、谷歌gemini-3.1-pro、Claude-4.6、文心ERNIE-X1.1、ERNIE-5.0、qwen3.6-max、qwen3.6-plus、百川、讯飞星火、商汤senseChat…

価格未確認AI・知識ベースClaude CodeOpenAI Agents使用前に確認
6163GitHub
AI・知識ベース

No visual example yet

Explore the skill

Langfuse

langfuse

🪢 Open source AI engineering platform: LLM evals, observability, metrics, prompt management, playground, datasets. Integrates with OpenTelemetry, LangChain, OpenAI SDK,…

価格未確認AI・知識ベースOpenAI AgentsLangChain使用前に確認
2.9万GitHub
開発・テスト

No visual example yet

Explore the skill

Opik

comet-ml

Debug, evaluate, and monitor your LLM applications, RAG systems, and agentic workflows with comprehensive tracing, automated evaluations, and production-ready dashboards.

価格未確認開発・テストLangChain使用前に確認
2万GitHub
AI・知識ベース

No visual example yet

Explore the skill

Agenta

Agenta-AI

The open-source LLMOps platform: prompt playground, prompt management, LLM evaluation, and LLM observability all in one place.

価格未確認AI・知識ベース使用前に確認
4450GitHub

ガイドと比較

開発者向け