暂未收录效果图
查看技能说明reproducibility
00200200
Audit or repair ML experiment reproducibility and review whether code changes preserve recorded outputs. Use for reproducibility blockers, before-and-after experiment ch…
OPENAGENTSKILL / DIRECTORY
为下一项任务找到合适的技能。探索适用于 Codex、Claude Code、Cursor 等 Agent 的工具。
9 Skills
搜索结果: 9
暂未收录效果图
查看技能说明00200200
Audit or repair ML experiment reproducibility and review whether code changes preserve recorded outputs. Use for reproducibility blockers, before-and-after experiment ch…
暂未收录效果图
查看技能说明m3dev
Gokart solves reproducibility, task dependencies, constraints of good code, and ease of use for Machine Learning Pipeline.
暂未收录效果图
查看技能说明claesbackman
Adversarially audit changed analysis code against a base ref, hunting for correctness errors in sample construction, merges, variable construction, silent failures, and…
暂未收录效果图
查看技能说明claesbackman
Review research code for reproducibility and quality, extract the paper's main empirical claims, compare paper to code, and write a constructive markdown report. Designe…
暂未收录效果图
查看技能说明Companion-Inc
Compare a paper's claims against its public codebase. Use when the user asks to audit a paper, check code-claim consistency, verify reproducibility of a specific paper,…
暂未收录效果图
查看技能说明ClawBio
Run read-only SQL against BigQuery public datasets with local result capture, cost safeguards, and reproducibility
暂未收录效果图
查看技能说明facebookresearch
BenchMARL is a library for benchmarking Multi-Agent Reinforcement Learning (MARL). BenchMARL allows to quickly compare different MARL algorithms, tasks, and models while…
暂未收录效果图
查看技能说明sweetcornna
Use when an award run needs external evidence — literature, datasets, benchmarks, domain constants, prior-art checks, or citation verification — with combined built-in W…
暂未收录效果图
查看技能说明alpacahq
Execute deterministic, reproducible historical backtests from a start date, end date, and strategy concept using the Alpaca CLI plus agent-written workspace code. Use wh…