No visual example yet
Explore the skillEvalscope
modelscope
A streamlined and customizable framework for efficient large model (LLM, VLM, AIGC) evaluation and performance benchmarking.
OPENAGENTSKILL / DIRECTORY
다음 작업에 맞는 스킬을 찾아보세요. Codex, Claude Code, Cursor 등을 지원합니다.
10 Skills
검색 결과: 10
No visual example yet
Explore the skillmodelscope
A streamlined and customizable framework for efficient large model (LLM, VLM, AIGC) evaluation and performance benchmarking.
No visual example yet
Explore the skillablab
No visual example yet
Explore the skillvotchallenge
No visual example yet
Explore the skillVoltAgent
A curated collection of AI agent research papers released in 2026, covering agent engineering, memory, evaluation, workflows, and autonomous systems.
No visual example yet
Explore the skillcertsocietegenerale
No visual example yet
Explore the skillalibaba
A CLI evaluation framework to make your Agent Skill Up.
No visual example yet
Explore the skillprobabl-ai
Track your Data Science. Skore's open-source Python library accelerates ML model development with automated evaluation reports, smart methodological guidance, and compre…
No visual example yet
Explore the skillApodexAI
Evaluation harness for Apodex-1.0 on public deep-research benchmarks.
No visual example yet
Explore the skillForward-Future
Practical repeatable AI-agent workflows for engineering, evaluation, operations, content, and design.
No visual example yet
Explore the skilltexttron
BrowseComp-Plus: A More Fair and Transparent Evaluation Benchmark of Deep-Research Agent (ACL 2026 Main)