No visual example yet
Explore the skillEvalscope
modelscope
A streamlined and customizable framework for efficient large model (LLM, VLM, AIGC) evaluation and performance benchmarking.
OPENAGENTSKILL / DIRECTORY
Trouvez un skill pour votre prochaine tâche avec Codex, Claude Code, Cursor et plus encore.
10 Skills
Résultats: 10
No visual example yet
Explore the skillmodelscope
A streamlined and customizable framework for efficient large model (LLM, VLM, AIGC) evaluation and performance benchmarking.
No visual example yet
Explore the skillablab
No visual example yet
Explore the skillvotchallenge
The official VOT Challenge evaluation and analysis toolkit
No visual example yet
Explore the skillVoltAgent
A curated collection of AI agent research papers released in 2026, covering agent engineering, memory, evaluation, workflows, and autonomous systems.
No visual example yet
Explore the skillcertsocietegenerale
No visual example yet
Explore the skillalibaba
A CLI evaluation framework to make your Agent Skill Up.
No visual example yet
Explore the skillprobabl-ai
Track your Data Science. Skore's open-source Python library accelerates ML model development with automated evaluation reports, smart methodological guidance, and compre…
No visual example yet
Explore the skillApodexAI
Evaluation harness for Apodex-1.0 on public deep-research benchmarks.
No visual example yet
Explore the skillForward-Future
Practical repeatable AI-agent workflows for engineering, evaluation, operations, content, and design.
No visual example yet
Explore the skilltexttron
BrowseComp-Plus: A More Fair and Transparent Evaluation Benchmark of Deep-Research Agent (ACL 2026 Main)