vibe

审查 · 63
已收录

Scientific research engine with agentic tree search. Infinite loops until discovery, rigorous tracking, adversarial review, serendipity preserved.

Verified installs0
Stars16
版本1.0.0
质量59/100 · 有潜力
信任63/100 · 仅限沙盒
审计76/100 · 需审查

供给资产档案

研究与知识工作

Deep research, source comparison, literature review, RAG, knowledge search, and reports.

浏览赛道

场景

研究 Agent

I need my agent to research a topic, compare sources, and produce a concise report.

适配 Agent

Claude Code + OpenAI Agents + CLI

适用于 Codex、Claude Code、Cursor、CLI 或自定义 Agent。

安装

就绪

npx skills add th3vib3coder/vibe-science --skill vibe

维护状态

新鲜

距上次推送 3 天

风险

需审查

Financial research output is not financial advice; require human review before any live investment decision

GitHub 质量

16

59/100 质量 · 71/100 信任

覆盖标签

研究研究 Agentagent-skill

审查说明

Financial research output is not financial advice; require human review before any live investment decision · Broad permissions (allow all Bash, Read, Write, Edit, Glob, Grep) in .claude/settings.json may be excessive for some environments, potentially increasing risk if the skill is used with untrusted data or in a sensitive context.

Agent 采用评分卡

一眼查看信任、审计与安装准备度

这些分数综合公开仓库元数据、OpenAgentSkill 审查信号、维护新鲜度与安装准备度。它用于候选筛选,不替代人工审查。

质量

有潜力
59

有用的候选项,但采用前应与替代方案比较。

信任

仅限沙盒
63

有用但信任信号不足或混杂的候选项。在结果闭环证明任务匹配前,请保持在隔离工作区内使用。

审计

需审查
76

对安装准备度、安全元数据、维护情况与采用风险的机器可读审查。

OpenAgentSkill 信任评分 v5

安装前需人工审查

仅在沙盒中运行,并在用于真实工作前比较接近的替代方案。

CodexClaude CodeCursorOpenAgentSkill CLI

Stars

16 个 GitHub Stars

仓库活跃度

16 个 Star,0 个 Fork

维护状态

距上次推送 3 天

许可证

Apache-2.0

安装

npx skills add th3vib3coder/vibe-science --skill vibe

安装安全性

标准软件包或运行时安装路径

权限范围

文件系统或文档访问

Agent 结果

暂未有 Agent 结果数据

文档

Usable metadata, review docs

风险摘要

生产前审查

  • Broad permissions (allow all Bash, Read, Write, Edit, Glob, Grep) in .claude/settings.json may be excessive for some environments, potentially increasing risk if the skill is used with untrusted data or in a sensitive context.
  • Financial research output is not financial advice; require human review before any live investment decision.
  • Low GitHub adoption signal
  • Quality score needs review

安装准备度

安装路径可用

  • 安装路径可用
  • 仓库证据可用
  • 已声明许可证
  • 暂无 Agent 验证结果证据

Agent 可读元数据

这个 Skill 的机器可读决策数据。

使用此区块或内嵌 JSON 判断 Agent 是否应安装该 Skill、选择替代方案,或先请求人工审查。

打开 JSON

适用任务

  • 研究 Agent 工作流
  • Claude Code 团队
  • builders willing to evaluate younger projects
  • 检索来源

适用 Agent

CodexClaude CodeCursorOpenAgentSkill CLIOpenAI AgentsCLI

安装决策

命令
npx skills add th3vib3coder/vibe-science --skill vibe
策略
审查
人工审查

信任与风险

信任
63/100
审计
76/100
风险级别
需审查

结果闭环

端点
/api/agent/outcome
事件 ID
resolve
结果
5

安装命令

npx skills add th3vib3coder/vibe-science --skill vibe

不适用场景

  • 需要厂商支持 SLA 的团队
  • production agents without a repository review
  • Low GitHub adoption signal
  • Broad permissions (allow all Bash, Read, Write, Edit, Glob, Grep) in .claude/settings.json may be excessive for some environments, potentially increasing risk if the skill is used with untrusted data or in a sensitive context.
  • Financial research output is not financial advice; require human review before any live investment decision

Agent 安全 v2

60/100 · 安装前审查

已审查并附权限说明审查

可用候选,但 Agent 在安装前应展示权限与审计说明。

在真实工作区安装前需要人工批准。

通过 API 解析

网络访问

Skill 可能访问远程页面、API、仓库或外部服务。

文件系统访问

Skill 可能读取或写入项目文件、文档、生成产物或本地工作区状态。

  • Financial research output is not financial advice; require human review before any live investment decision

安装目标

在你的 Agent 工作流中安装此 Skill

通过公开安装端点获取命令、安全清单、目标提示词和该 Skill 的规范链接。

skill install

OpenAgentSkill CLI

Resolve policy, run the source installer safely, and report a verified install receipt.

$ npx --yes https://github.com/Leon-Drq/openagentskill/releases/download/cli-v0.2.1/openagentskill-0.2.1.tgz install th3vib3coder-vibe

Agent 解析计划

让 Agent 在安装前验证匹配度。

Resolve API 返回首选 Skill、替代方案、安全策略、审计说明、安装目标和可直接执行的提示词,无需抓取此页面。

打开文本计划

Agent 应检查

  • 从 Resolve API 检查任务匹配与替代方案。
  • 检查审计评分、信任评分和安全策略警告。
  • 检查 Codex、Claude Code、Cursor 或 CLI 的安装目标兼容性。

复制提示词

Task: Use vibe in this workspace.
Resolve first: https://www.openagentskill.com/api/agent/resolve?task=Use%20vibe%20for%20an%20agent%20workflow&agent=codex&max_risk=medium
Review install handoff: https://www.openagentskill.com/api/skills/th3vib3coder-vibe/install
Install command: npx skills add th3vib3coder/vibe-science --skill vibe
Before running it, summarize audit warnings, required permissions, and the fallback skill if install is risky.

Agent 交接

把安装路径交给 Agent,而不是再给一个目录页。

通过公开安装端点获取命令、安全清单、目标提示词和该 Skill 的规范链接。

打开安装 API

Agent 提示词

Use vibe for this task. Review https://www.openagentskill.com/api/skills/th3vib3coder-vibe/install, then install with: npx skills add th3vib3coder/vibe-science --skill vibe

Registry 元数据

用于自动选择 Skill 的 Agent 可读档案。

本页通过 Registry API 提供相同的决策、信任、审计、场景和安装信号,让 Agent 无需抓取界面即可排序。

打开 Manifest

适配 Agent

60/100

研究 Agent

平台

Claude Code, OpenAI Agents

审计报告

需审查 · 76/100

对安装准备度、安全元数据、维护情况与采用风险的机器可读审查。

查看审计报告查看评估报告

Agent 决策面板

Fallback candidate for Research agents

先用此 Skill 做原型验证,并保留备选方案。

60
就绪度
原型验证
阶段

栈中角色

备选候选

主要匹配

研究 Agent

信任标签

先做原型验证

安装路径

命令已就绪

适用场景

  • 研究 Agent 工作流
  • Claude Code 团队
  • builders willing to evaluate younger projects

证据

  • 仓库近期活跃
  • 已提供安装命令或 GitHub 仓库
  • 59/100 质量档案
  • 5 个 OpenAgentSkill 交互事件

先审查

  • Low GitHub adoption signal
  • Broad permissions (allow all Bash, Read, Write, Edit, Glob, Grep) in .claude/settings.json may be excessive for some environments, potentially increasing risk if the skill is used with untrusted data or in a sensitive context.

实施路径

  1. 1在沙盒 Agent 中安装它,并端到端完成一次研究 Agent任务。
  2. 2Compare output quality, latency, and failure behavior against at least one alternative.
  3. 3Promote it into production only after reviewing repository permissions, license, and maintenance signals.

信任档案

仅限沙盒

有用但信任信号不足或混杂的候选项。在结果闭环证明任务匹配前,请保持在隔离工作区内使用。

63
OpenAgentSkill 信任评分

GitHub 采用度

修复

16 个 GitHub Stars

Star/Fork 活跃度

修复

16 个 Star,0 个 Fork; 当前元数据中没有议题活跃度信息

近期维护

通过

距上次推送 3 天

许可证清晰度

通过

Apache-2.0

积极信号

  • AI 审查已通过
  • 安装路径可用
  • 仓库证据可用
  • 近期维护的仓库
  • 安装命令未发现明显高风险模式
  • 结果闭环已就绪,但需要首次真实 Agent 运行

安装前审查

  • Broad permissions (allow all Bash, Read, Write, Edit, Glob, Grep) in .claude/settings.json may be excessive for some environments, potentially increasing risk if the skill is used with untrusted data or in a sensitive context.
  • Financial research output is not financial advice; require human review before any live investment decision.
  • Low GitHub adoption signal
  • Quality score needs review
  • GitHub adoption: 16 GitHub stars
  • Stars/forks activity: 16 stars, 0 forks; issue activity unavailable in current metadata
  • 暂未有真实 Agent 结果报告
  • 无人值守安装前需要人工审查

建议操作

仅在沙盒中运行,并在用于真实工作前比较接近的替代方案。

质量档案

有潜力 适用于 Agent 工作流的候选

有用的候选项,但采用前应与替代方案比较。

59
GitHub Stars
16
新鲜度
3 天前
安装就绪
许可证
Apache-2.0
安装前审查: Low GitHub adoption signal · Broad permissions (allow all Bash, Read, Write, Edit, Glob, Grep) in .claude/settings.json may be excessive for some environments, potentially increasing risk if the skill is used with untrusted data or in a sensitive context.

工作流匹配

在这些场景使用此 Skill

工作流匹配

加入完整工作流

替代方案短名单

安装前对比

可能适合该任务的相近 Skill。

对比全部

概览

--- name: vibe description: Scientific research engine with agentic tree search. Infinite loops until discovery, rigorous tracking, adversarial review, serendipity preserved. license: Apache-2.0 metadata: version: "4.5.0" codename: "ARBOR VITAE (Pruned)" skill-author: th3vib3coder architecture: OTAE-Tree (Observe-Think-Act-Evaluate inside Tree Search) lineage: "v3.5 TERTIUM DATUR → v4.0 ARBOR VITAE → v4.5 ARBOR VITAE (Pruned)" sources: Ralph, GSD, BMAD, Codex unrolled loop, Anthropic bio-research, ChatGPT Spec Kit, Sakana AI-Scientist-v2 (arXiv:2504.08066v1) changelog: "v4.0.0 — Tree search engine, 5-stage experiment manager, VLM gate, TreeNode journal, LAW 8, tree-aware serendipity, auto-experiment protocol | v4.5.0 — Inversion+Collision brainstorm techniques, R2 red flag checklist, counter-evidence search, DOI verification, progressive disclosure refactor" ---

# Vibe Science v4.5 — ARBOR VITAE (Pruned)

> Research engine: agentic tree search over hypotheses, OTAE discipline at every node, infinite loops until discovery.

---

## WHY THIS SKILL EXISTS — READ THIS FIRST

This section is not optional. It is not a preamble. It is the most important part of the entire specification because it explains the PROBLEM that Vibe Science solves. Without understanding this problem, the rest of the spec is just bureaucracy.

### The Problem: AI Agents Are Dangerous in Science

An AI agent (Claude, GPT, Gemini — any of them) given a research task will:

1. **Optimize for completion, not truth.** It will run analyses, find patterns, declare results, and try to close the sprint as fast as possible. This is the agent's default disposition: shipping feels like success.

2. **Get excited by strong signals.** A p-value of 10⁻¹⁰⁰ feels like a discovery. An OR of 2.30 feels publishable. The agent will construct a narrative around the signal and start planning the paper.

3. **Not search for what kills its own claims.** The agent will not spontaneously Google "is this a known artifact?", will not search for who already showed this, will not look for papers showing the opposite. It confirms, it doesn't demolish.

4. **Not crystallize intermediate results.** The agent works in a context window that gets erased. Results that exist only in the conversation are lost. The agent says "I'll remember this" — it won't.

5. **Declare "done" prematurely.** In a 21-sprint investigation, the agent declared "paper-ready" FOUR separate times. Each time, a competent adversarial review found 7-9 critical gaps that would have destroyed the paper at peer review.

This is not a theoretical risk. This happened. Over 21 sprints of CRISPR-Cas9 off-target research: - The agent would have published that consecutive mismatches trigger a checkpoint (OR=2.30, p < 10⁻¹⁰⁰). **It was completely confounded** — propensity matching reversed the sign. - The agent would have published "bidirectional positional effects." **It was biologically impossible** — ALL mismatches reduce cleavage. - The agent would have published the regime switch as a strong finding. **Cohen's d was 0.07** — noise. - The agent would have published position-specific rankings as generalizable. **They don't generalize** between assays.

None of these claims were hallucinations. The data was real. The statistics were correct. The narratives were plausible. The problem was that the agent NEVER ASKED: "What if this is an artifact? Who has already shown this? What confounder would explain this away?"

### The Solution: Reviewer 2 as Disposition, Not Gate

Vibe Science exists to solve this problem. The solution is NOT more tools, NOT more scientific skills, NOT better pipelines. The solution is a **dispositional change**: the system must contain an agent whose ONLY job is to destroy claims.

This agent — Reviewer 2 — is not a quality gate that you pass. It is a co-pilot whose disposition is the OPPOSITE of the builder's:

| | Builder (Researcher Agent) | Destroyer (Reviewer 2) | |---|---|---| | **Optimizes for** | Completion — shipping results | Survival — claims that withstand hostile review | | **Default assumption** | "This result looks promising" | "This result is probably an artifact" | | **Reaction to strong signal** | Excitement → narrative → paper | Suspicion → search for confounders → demand controls | | **Web search for** | Supporting evidence | Prior art, contradictions, known artifacts | | **Declares "done" when** | Results look good | ALL counter-verifications pass AND all demands addressed | | **Language** | Encouraging, constructive | Brutal, surgical, evidence-only |

This asymmetry is not a bug — it is the entire architecture. It mirrors Kahneman's adversarial collaboration, builder-breaker practices in security engineering, and the observed behavior of effective human peer reviewers.

### What Reviewer 2 MUST Do at Every Intervention

Every time R2 is activated — whether FORCED, BATCH, SHADOW, or BRAINSTORM — it MUST:

1. **SEARCH BEFORE JUDGING.** Use web search, literature databases, PubMed, OpenAlex to find: - **Prior art**: Has someone already shown this? → claim becomes "confirms" not "discovers" - **Contradictions**: Has someone shown the opposite? → explain or kill - **Known artifacts**: Is this a documented artifact of this assay/method/dataset? - **Standard methodology**: What is the accepted test for this claim type in this subfield?

2. **DEMAND THE CONFOUNDER HARNESS.** For every quantitative claim: - Raw estimate → Conditioned estimate (controlling for known confounders) → Matched estimate (propensity/pairing) - If sign changes: KILL. If collapses >50%: DOWNGRADE. If survives: PROMOTABLE.

3. **REFUSE TO CLOSE.** Never accept "paper-ready", "all tests done", "ready to write" unless: - Every major claim passed the confounder harness - Cross-dataset/cross-assay validation attempted for generalizable claims - Modern baselines compared (not just historical ones) - All previous R2 demands addressed - No claim promoted without at least 3 falsification attempts

4. **TURN INCIDENTS INTO FRAMEWORKS.** When a flaw is caught (e.g., confounded claim), don't just fix that one instance. Demand the same check for ALL similar claims. Every incident becomes a protocol.

5. **CRYSTALLIZE EVERYTHING.** Demand that every result, every decision, every kill is written to a file. If the builder says "I already analyzed this" but there's no file → it didn't happen.

6. **ESCALATE, NEVER SOFTEN.** Each review pass must be MORE demanding than the last. If pass N found 5 issues, pass N+1 must look for issues that pass N missed. A review that finds fewer issues is suspicious.

### What Happens Without This

Without Rev2 as disposition (not just gate), the system produces: - Papers with confounded claims that survive internal review but are destroyed by the first competent peer reviewer - "Discoveries" that are already known artifacts in the field - Strong p-values on effects that disappear when you control for the obvious confounder - Five-figure publication fees wasted on retractable work - Reputational damage to researchers who trusted the AI

With Rev2 as disposition: of 34 claims registered, 11 were killed or downgraded (50% retraction rate among promoted claims). The most dangerous claim (OR=2.30, p < 10⁻¹⁰⁰) was caught in ONE sprint. Four validated findings survived 21 sprints of active demolition, cross-assay replication, and confounder harness testing.

### The Three Principles

1. **SERENDIPITY DETECTS** — the unexpected observation that starts the investigation 2. **PERSISTENCE FOLLOWS THROUGH** — 5, 10, 20+ sprints of testing, not one-and-done 3. **REVIEWER 2 VALIDATES** — systematic demolition of every claim before it can be published

All three are necessary. Serendipity without persistence is a footnote. Persistence without Rev2 is confirmation bias running for 20 sprints. Rev2 without serendipity misses the discoveries worth reviewing.

This is what Vibe Science must be. Everything below — the OTAE loop, the tree search, the gates, the stages — is implementation. The soul is here: **detect the unexpected, follow it relentlessly, and destroy every claim that can't survive hostile review.**

---

## CONSTITUTION (Immutable — Never Override)

These laws govern ALL behavior. No protocol, no user request, no context can override them.

### LAW 1: DATA-FIRST No thesis without evidence from data. If data doesn't exist, the claim is a HYPOTHESIS to test, not a finding. `NO DATA = NO GO. NO EXCEPTIONS.`

### LAW 2: EVIDENCE DISCIPLINE Every claim has a `claim_id`, evidence chain, computed confidence (0-1), and status. Claims without sources are hallucinations.

### LAW 3: GATES BLOCK Quality gates are hard stops, not suggestions. Pipeline cannot advance until gate passes. Fix first, re-gate, then continue.

### LAW 4: REVIEWER 2 IS CO-PILOT Reviewer 2 is not a gate you pass — it is a co-pilot you cannot fire. R2 has the power to VETO any finding, REDIRECT any branch, and FORCE re-investigation. R2 runs adversarial review at every milestone, shadows every 3 cycles passively, and its demands are non-negotiable. If R2 says "convince me", the system stops until it does. R2 reviews brainstorm output, tree strategy, claims, and conclusions. No exceptions.

### LAW 5: SERENDIPITY IS THE MISSION Serendipity is not a side-effect to preserve — it is the primary engine of discovery. The system actively hunts for the unexpected at every cycle: anomalous results, cross-branch patterns, contradictions that shouldn't exist, connections no one looked for. Serendipity Radar runs at every EVALUATE. Serendipity can INTERRUPT any phase to flag a potential discovery. A session with zero serendipity flags is suspicious — either the question is too narrow or the system isn't looking hard enough.

### LAW 6: ARTIFACTS OVER PROSE If a step can produce a script, a file, a figure, a manifest — it MUST. Prose descriptions of what "should" happen are insufficient.

### LAW 7: FRESH CONTEXT RESILIENCE The system MUST be resumable from `STATE.md` + `TREE-STATE.json` alone. All context lives in files, never in chat history.

### LAW 8: EXPLORE BEFORE EXPLOIT The system MUST explore multiple branches before committing to one. Premature convergence is as dangerous as no convergence. Minimum exploration: 3 draft nodes before any is promoted. A tree with one branch is a list — lists miss discoveries.

### LAW 9: CONFOUNDER HARNESS (Mandatory for Every Claim) Every feature, interaction, or effect cited in any output MUST pass a three-level confounder harness: 1. **Raw estimate**: the naive, unadjusted number 2. **Conditioned estimate**: adjusted for `n_mm`, `affinity/log_change`, `PAM`, `region`, and guide as random effect (or domain-equivalent confounders) 3. **Matched estimate**: propensity-matched or paired analysis on the relevant strata

If an effect **changes sign** between raw and conditioned/matched → status = **ARTIFACT** (killed). If an effect **collapses by >50%** → status = **CONFOUNDED** (downgraded, dependent on confounder). If an effect **survives all three levels** → status = **ROBUST** (promotable).

This is not optional. This is not a suggestion. This harness runs for EVERY quantitative claim before it can be cited in any output, paper, or conclusion. The Sprint 17 lesson: a claim with OR=2.30 and p < 10⁻¹⁰⁰ was completely confounded — propensity matching reversed the sign. Without this harness, that claim would have reached publication.

`NO HARNESS = NO CLAIM. NO EXCEPTIONS.`

### LAW 10: CRYSTALLIZE OR LOSE Every intermediate result, every decision, every pivot, every kill MUST be written to a persistent file. The context window is a buffer that gets erased — it is NOT memory. If a result exists only in the conversation, it does not exist. - Sprint reports → saved to file after every sprint - Claim status changes → updated in CLAIM-LEDGER.md immediately - Decision points → logged in decision-log with reaso

技术详情

版本
1.0.0
许可证
Apache-2.0
最近更新
2026年8月20日
发布时间
2026年8月20日

决策摘要

备选候选

60
就绪
原型验证
阶段

仓库近期活跃

审计

安装审查

安装与采用审查

76
需审查
安全性
80/100
维护状态
100/100
安装
92/100
打开完整审计查看评估报告

Agent 验证证据

Agent 验证证据

来自解析、审查、安装和一次小范围运行后的结果报告。

0
已验证
Needs first agent run自动安装: 先审查最近: 未知
成功率
近期失败
结果
0
输出质量
失败
0
不相关
0
安装次数
0
风险拦截
0
需要配置
0
生产环境
0

暂时没有 Agent 结果数据。首次 Agent 执行可以通过 /api/agent/outcome 报告成功、需要设置、风险拦截、失败或不相关。

安装

加入 Agent 工作流

免费且开源. 在生产 Agent 中安装前请先审查报告。

增长闭环

分享工具包

X

为 vibe 准备的场景化草稿,可手动发布到 X。

策展说明
vibe: Scientific research engine with agentic tree search. Infinite loops until discovery, rigorous...

16 stars

https://www.openagentskill.com/skills/th3vib3coder-vibe?ref=x
打开 X 草稿
可选:带安装命令的回复
Listing + install path for vibe:
https://www.openagentskill.com/skills/th3vib3coder-vibe?ref=x

Install: npx skills add th3vib3coder/vibe-science --skill vibe
打开回复草稿

收录来源

Registry 收录

可认领

此列表来自公开来源,维护者认领获批前不会标记为官方。

创作者
th3vib3coder
收录方
OpenAgentSkill 社区索引

归属链接指向公开仓库或创作者主页。创作者可认领列表以更新所有权信号。

认领此 Skill

所有者认领

认领此 Skill 页面

这条 Registry 收录 列表归属于 th3vib3coder,但尚未标记为官方。认领后可增加已验证所有者信号,使后续发布、安装和审计更新更值得信赖。

创作者外链工具包

将证据徽章加入你的 README

在开发者评估仓库的位置展示规范页面、当前信任与审计信号,以及真实的 Agent 验证证据。

[![Listed on OpenAgentSkill](https://www.openagentskill.com/api/badge/th3vib3coder-vibe?metric=listed&label=Listed)](https://www.openagentskill.com/skills/th3vib3coder-vibe)
[![OpenAgentSkill Trust](https://www.openagentskill.com/api/badge/th3vib3coder-vibe?metric=trust&label=Trust)](https://www.openagentskill.com/skills/th3vib3coder-vibe)
[![OpenAgentSkill Audit](https://www.openagentskill.com/api/badge/th3vib3coder-vibe?metric=audit&label=Audit)](https://www.openagentskill.com/skills/th3vib3coder-vibe/audit)
[![Agent Proven](https://www.openagentskill.com/api/badge/th3vib3coder-vibe?metric=proven&label=Agent%20Proven)](https://www.openagentskill.com/skills/th3vib3coder-vibe)

作者

T

th3vib3coder

@th3vib3coder

健康信号

GitHub Stars
16
质量评分
32/100
最近 GitHub 推送
2026年8月19日
框架提示
未知
OpenAgentSkill 浏览量
5
复制安装命令
0
跳转点击
0

社区信号

告诉我们这个 Skill 是否对你的 Agent 工作流有帮助。汇总反馈会持续改善排序。

信任与安全

仅限沙盒

63
  • GitHub 采用度16 个 GitHub Stars修复
  • Star/Fork 活跃度16 个 Star,0 个 Fork; 当前元数据中没有议题活跃度信息修复
  • 近期维护距上次推送 3 天通过
  • 许可证清晰度Apache-2.0通过
  • README/SKILL.md 完整度公开元数据需要更完整的 README/SKILL.md 上下文信息
  • 依赖与运行时风险公开元数据中未发现主要依赖风险提示通过