OPENAGENTSKILL / DIRECTORY

AI Agent Skills

为下一项任务找到合适的技能。探索适用于 Codex、Claude Code、Cursor 等 Agent 的工具。

搜索结果 · “benchmarking”

44 Skills

搜索结果: 44

设计与 UI

暂未收录效果图

查看技能说明

Design and validity review for studies that benchmark one or more AI systems against a human-expert panel as the reference. Covers the evaluation question and arm defini…

价格未确认设计与 UIClaude Code
291GitHub
AI 与知识库

暂未收录效果图

查看技能说明

Bocoel

rentruewang

Bayesian Optimization as a Coverage Tool for Evaluating LLMs. Accurate evaluation (benchmarking) that's 10 times faster with just a few lines of modular code.

价格未确认AI 与知识库使用前请审核
289GitHub
AI 与知识库

暂未收录效果图

查看技能说明

Agentops

AgentOps-AI

Python SDK for AI agent monitoring, LLM cost tracking, benchmarking, and more. Integrates with most LLMs and agent frameworks including CrewAI, Agno, OpenAI Agents SDK,…

价格未确认AI 与知识库Claude CodeOpenAI Agents使用前请审核
5759GitHub
视频与音频

暂未收录效果图

查看技能说明

Video ChatGPT

mbzuai-oryx

[ACL 2024 🔥] Video-ChatGPT is a video conversation model capable of generating meaningful conversation about videos. It combines the capabilities of LLMs with a pretrai…

价格未确认视频与音频OpenAI Agents使用前请审核
1504GitHub
开发与测试

暂未收录效果图

查看技能说明

Golang benchmarking, profiling, and performance measurement. Use when writing, running, or comparing Go benchmarks, profiling hot paths with pprof, interpreting CPU/memo…

价格未确认开发与测试Claude Code
3010GitHub

指南与对比

开发者入口