Deep Research Bench
Ayanami0730
DeepResearch Bench: A Comprehensive Benchmark for Deep Research Agents
OPENAGENTSKILL / DIRECTORY
Find a skill for your next task. Explore tools for Codex, Claude Code, Cursor and more.
1–7 / 7
Results: 7
Ayanami0730
DeepResearch Bench: A Comprehensive Benchmark for Deep Research Agents
onyx-dot-app
Dataset and benchmark for RAG on company internal documents.
GaeaRuiW
GaeaRuiW/kube-llmops is a high-star GitHub project relevant to AI agent workflows.
K-Dense-AI
Autonomously improve a real artifact (code, training recipe, agent harness, data pipeline, prompt) against an objective and an evaluator, using Hypothesis Tree Refinemen…
google-ai-edge
Shrink a converted LiteRT model with ai-edge-quantizer (fp16 / int8 / int4) without losing accuracy, verifying parity against the float source after every step. Use when…
Impertio-Studio
Use when reviewing or validating Frappe/ERPNext code against best practices and common pitfalls. Checks generated code before deployment, validates against all 61 frappe…
Impertio-Studio
Use when implementing translations/i18n in Frappe v14-v16 apps. Covers _() in Python, __() in JavaScript, CSV translation files, bench commands, string extraction rules,…
Owner-curated external sources. Not filtered by the scores or compatibility controls above; excluded from GitHub rankings and automatic installation.
No external entries match this query.