OPENAGENTSKILL / DIRECTORY

AI Agent Skills

为下一项任务找到合适的技能。探索适用于 Codex、Claude Code、Cursor 等 Agent 的工具。

搜索结果 · “llm-evaluation”

109 Skills

搜索结果: 109

AI 与知识库

暂未收录效果图

查看技能说明

非线智能 NoneLinear - ReLE评测:中文AI大模型能力评测(持续更新):目前已囊括374个大模型,覆盖chatgpt、gpt-5.4、谷歌gemini-3.1-pro、Claude-4.6、文心ERNIE-X1.1、ERNIE-5.0、qwen3.6-max、qwen3.6-plus、百川、讯飞星火、商汤senseChat…

价格未确认AI 与知识库Claude CodeOpenAI Agents使用前请审核
6163GitHub
AI 与知识库

暂未收录效果图

查看技能说明

Langfuse

langfuse

🪢 Open source AI engineering platform: LLM evals, observability, metrics, prompt management, playground, datasets. Integrates with OpenTelemetry, LangChain, OpenAI SDK,…

价格未确认AI 与知识库OpenAI AgentsLangChain使用前请审核
2.9万GitHub
开发与测试

暂未收录效果图

查看技能说明

Opik

comet-ml

Debug, evaluate, and monitor your LLM applications, RAG systems, and agentic workflows with comprehensive tracing, automated evaluations, and production-ready dashboards.

价格未确认开发与测试LangChain使用前请审核
2万GitHub
AI 与知识库

暂未收录效果图

查看技能说明

Agenta

Agenta-AI

The open-source LLMOps platform: prompt playground, prompt management, LLM evaluation, and LLM observability all in one place.

价格未确认AI 与知识库使用前请审核
4450GitHub

指南与对比

开发者入口