Codebase-level adversarial review by a panel of frontier models. A Claude Code skill that runs every file through DeepSeek + Gemini + Kimi + MiniMax in sequence, then has Claude verify the findings against the actual source.
Supply asset profile
Code review, repo analysis, testing, CI, GitHub, DevOps, and developer workflow skills.
Scenario
Coding agents
I need a coding agent that can understand a repository, edit code, and review pull requests.
Agent fit
Claude Code + CLI + Codex
Codex, Claude Code, Cursor, CLI, or custom agents.
Install
Ready
npx skills add Bambushu/crucible
Maintenance
fresh
1d since push
Risk
Needs review
Permission surface may require sandboxing
GitHub quality
48
75/100 quality · 74/100 trust
Coverage tags
Review notes
Permission surface may require sandboxing · Low GitHub adoption signal
Agent adoption scorecard
These scores combine public repository metadata, OpenAgentSkill review signals, maintenance freshness, and install readiness. They are a shortlist signal, not a replacement for human review.
Quality
StrongSolid option that is likely worth shortlisting for production workflows.
Trust
Sandbox onlyUseful candidate with missing or mixed trust signals. Keep it in an isolated workspace until the outcome loop proves task fit.
Audit
Needs reviewInstall readiness, security metadata, maintenance, and adoption risk.
Trust Score v5
Run only in a sandbox and compare close alternatives before using it for real work.
Stars
48 GitHub stars
Repo activity
48 stars, 6 forks
Maintenance
1d since push
License
MIT
Install
npx skills add Bambushu/crucible
Install safety
standard package or runtime install path
Permission surface
secrets or environment access, shell or command execution
Agent outcomes
No agent outcome data yet
Docs
Strong README/SKILL.md context
Risk summary
Install readiness
Agent-readable metadata
Use this block or the embedded JSON to decide whether an agent should install this skill, choose an alternative, or ask for human review first.
Suited tasks
Suited agents
Install decision
Trust and risk
Outcome loop
Do not use when
Alternative
164.7K stars
npx skills add mattpocock/skills --skill grill-with-docs
Alternative
168.6K stars
npx skills add mattpocock/skills --skill code-review
Alternative
164.7K stars
npx skills add mattpocock/skills --skill to-spec
Alternative
176.7K stars
npx skills add mattpocock/skills --skill to-tickets
Agent safety v2
This skill should not be selected by an agent without explicit human security review.
Do not auto-install. Inspect the source, dependencies, and permission surface first.
high
Skill metadata references terminal, CLI, shell, subprocess, or command execution workflows.
medium
Skill likely fetches remote pages, APIs, repositories, or external services.
medium
Skill may read or write project files, documents, generated artifacts, or local workspace state.
high
Skill metadata references credentials, tokens, environment variables, or secret-bearing workflows.
Install targets
Copy the registry command or an agent-specific install prompt for Codex, Claude Code, and Cursor.
Use the registry command when your workflow supports the OpenAgentSkill installer.
$ npx skills add Bambushu/crucibleAgent resolve plan
The Resolve API returns the selected skill, alternatives, safety policy, audit notes, install target, and copy-paste prompt an agent can follow without scraping this page.
Resolve JSON
/api/agent/resolve?task=Use%20Crucible%20for%20an%20agent%20workflow&agent=codex&max_risk=medium
Resolve text
/api/agent/resolve?task=Use%20Crucible%20for%20an%20agent%20workflow&agent=codex&max_risk=medium&format=text
Install handoff
/api/skills/bambushu-crucible/install
Agent should check
Copy prompt
Task: Use Crucible in this workspace.
Resolve first: https://www.openagentskill.com/api/agent/resolve?task=Use%20Crucible%20for%20an%20agent%20workflow&agent=codex&max_risk=medium
Review install handoff: https://www.openagentskill.com/api/skills/bambushu-crucible/install
Install command: npx skills add Bambushu/crucible
Before running it, summarize audit warnings, required permissions, and the fallback skill if install is risky.Agent handoff
Use the public install endpoint to fetch the command, safety checklist, target prompts, and canonical links for this skill.
Install handoff
/api/skills/bambushu-crucible/install
LLM text format
/api/skills/bambushu-crucible/install?format=text
Find alternatives
/api/skills/search?q=Crucible&limit=3
Agent prompt
Use Crucible for this task. Review https://www.openagentskill.com/api/skills/bambushu-crucible/install, then install with: npx skills add Bambushu/crucibleRegistry metadata
This page exposes the same decision, trust, audit, use-case, and install signals through the Registry API, so agents can rank this skill without scraping the UI.
Manifest
/api/registry/manifest/bambushu-crucible
LLM text
/api/registry/manifest/bambushu-crucible?format=text
Install alias
/api/registry/install/bambushu-crucible
Recommend
/api/registry/recommend?task=Use%20Crucible%20in%20an%20agent%20workflow&limit=3
Agent fit
Coding agents
Use-case tags
Platforms
Python, Claude Code
Audit report
Review install readiness, maintenance, trust, quality, and metadata warnings before adding this skill to an agent workflow.
Agent decision cockpit
Shortlist this skill and compare it with close alternatives before production adoption.
Role in stack
Companion skill
Primary fit
Coding agents
Trust label
Strong shortlist
Install path
Command ready
Use when
Evidence
Review first
Implementation path
Trust profile
Useful candidate with missing or mixed trust signals. Keep it in an isolated workspace until the outcome loop proves task fit.
GitHub adoption
CHECK48 GitHub stars
Stars/forks activity
CHECK48 stars, 6 forks; issue activity unavailable in current metadata
Recent maintenance
PASS1d since push
License clarity
PASSMIT
Good signals
Review before install
Recommended action
Run only in a sandbox and compare close alternatives before using it for real work.
Quality profile
Solid option that is likely worth shortlisting for production workflows.
Workflow fit
Build and ship code
I need a coding agent that can understand a repository, edit code, and review pull requests.
Manage repositories
I need my agent to triage GitHub issues, review pull requests, and summarize repository changes.
Investigate faster
I need my agent to research a topic, compare sources, and produce a concise report.
Stack fit
Inspect, patch, and verify code
A stack for software agents that inspect repositories, review pull requests, generate tests, and turn findings into shippable patches.
Find, compare, and synthesize
A stack for agents that gather sources, compare claims, summarize long material, and draft useful research briefs.
Operate and verify web apps
A stack for agents that navigate products, fill forms, take screenshots, and verify real user flows across web applications.
Alternative shortlist
Similar skills in this category, ranked with the same readiness and quality signals.
A relentless interview that pressure-tests a plan against the codebase, sharpens domain language, and updates CONTEXT.md and ADRs when decisions become durable.
Review a branch or diff against repository standards and the originating spec in two independent analysis passes.
Turn the current conversation and codebase context into a structured implementation spec, then publish it to the configured project issue tracker.
Break a plan, spec, or conversation into independently actionable tracer-bullet tickets with explicit blocking relationships.
<div align="center">
<img src="assets/hero.png" alt="Crucible" width="520" />
# Crucible
**Codebase-level adversarial review by a panel of frontier models.**
A Claude Code skill that walks your code piece-by-piece and puts every file under simultaneous pressure from a panel of structurally different models, then aggregates the findings into a single severity-ranked report that Claude itself verifies before you see it.
[Install](#install) · [How it works](#how-it-works) · [Cost](#cost) · [Modes](#modes) · [Sample report](#sample-report)
</div>
---
## What it is
`/crucible` is a [Claude Code](https://claude.com/claude-code) slash-command skill. You drop the folder into `~/.claude/skills/`, set one env var, and from inside any project you can run:
``` /crucible # review the current branch's diff /crucible --all # review the whole repo /crucible --paths "src/api/**/*.ts" # review a glob /crucible --diff main...HEAD # review a specific range /crucible --verify # also EXECUTE a repro for runtime-flagged findings ```
Behind the scenes, Claude:
1. Resolves the file list and prints a pre-flight (files, models, est. cost). 2. Loads a panel of four current SOTA paid models from OpenRouter, each from a different vendor family (e.g. DeepSeek, Google, Moonshot, MiniMax). 3. Reviews every in-scope file through the panel: pass 1 finds, pass 2 validates, pass 3 consolidates with severity ranks. 4. Runs one cross-file architectural meta-pass to catch repeated anti-patterns, missing layers, and coupling smells. 5. **(opt-in `--verify`) Executes a repro for the runtime bugs.** Some bugs only exist when the code runs — a retry the surrounding rate-limit always rejects, a `2>/dev/null` that hides a device error. For each finding the panel tags *runtime-checkable*, a model writes a minimal repro harness, Crucible runs it in a **locked sandbox**, and findings that actually reproduce are promote
Frameworks & Tools
Decision snapshot
recent repository activity
Audit snapshot
Install and adoption review
Agent-proven evidence
Outcome reports after resolve, review, install, and one narrow run.
No agent outcome data yet. The first agent run can report success, setup needs, risk blocks, failure, or not-relevant through /api/agent/outcome.
Install
Free and open source. Review the audit before production use.
Growth loop
Scenario-led draft for Crucible, ready for a manual X post.
Most coding agents don't fail from lack of model power. They fail when repo context disappears. Crucible gives coding agents a repeatable way to plan, patch, review, or ship. 48 stars https://www.openagentskill.com/skills/bambushu-crucible?ref=x #AIAgents
Listing + install path for Crucible: https://www.openagentskill.com/skills/bambushu-crucible?ref=x Install: npx skills add Bambushu/crucible
Listing source
This listing was indexed from public sources and is not marked official until a maintainer claim is approved.
Attribution links to the public repository or creator profile. Creators can claim the listing to update ownership signals.
Claim this skillOwner claim
This community indexed listing is attributed to Bambushu but is not marked official yet. Claim it to add a verified owner signal and make future launch, install, and audit updates easier to trust.
Creator backlink kit
Show the canonical listing, current trust and audit signals, and real Agent Proven evidence where developers evaluate the repository.
[](https://www.openagentskill.com/skills/bambushu-crucible)
[](https://www.openagentskill.com/skills/bambushu-crucible)
[](https://www.openagentskill.com/skills/bambushu-crucible/audit)
[](https://www.openagentskill.com/skills/bambushu-crucible)Bambushu
@bambushu
Share whether this skill looks useful for your agent workflow. Aggregated feedback improves rankings over time.
Sandbox only
Grill With Docs
A relentless interview that pressure-tests a plan against the codebase, sharpens domain language, and updates CONTEXT.md and ADRs when decisions become durable.
164.7K stars · 0 installsCode Review
Review a branch or diff against repository standards and the originating spec in two independent analysis passes.
168.6K stars · 0 installsTo Spec
Turn the current conversation and codebase context into a structured implementation spec, then publish it to the configured project issue tracker.
164.7K stars · 0 installsTo Tickets
Break a plan, spec, or conversation into independently actionable tracer-bullet tickets with explicit blocking relationships.
176.7K stars · 0 installs