Claude Code Vision Skill
为 Claude Code 赋能多模态视觉能力,支持豆包、通义千问、GPT-4o 等模型,用于截图 / UI / 图表分析;适配 DeepSeek 等无视觉底座,搭配 browser-harness 可做前端布局自动化检查。
Supply asset profile
Coding and developer agents
Code review, repo analysis, testing, CI, GitHub, DevOps, and developer workflow skills.
Scenario
Coding agents
I need a coding agent that can understand a repository, edit code, and review pull requests.
Agent fit
Claude Code + OpenAI Agents + Browser agents
Codex, Claude Code, Cursor, CLI, or custom agents.
Install
Ready
npx skills add xiincs/claude-code-vision-skill
Maintenance
fresh
1d since push
Risk
Needs review
Dependency or permission surface needs review
GitHub quality
92
77/100 Quality · 72/100 Trust
Coverage tags
Review notes
Dependency or permission surface needs review · Permission surface may require sandboxing
Agent adoption scorecard
Trust, audit, and install readiness at a glance
These scores combine public repository metadata, OpenAgentSkill review signals, maintenance freshness, and install readiness. They are a shortlist signal, not a replacement for human review.
Quality
StrongSolid option that is likely worth shortlisting for production workflows.
Trust
Sandbox onlyUseful candidate with missing or mixed trust signals. Keep it in an isolated workspace until the outcome loop proves task fit.
Audit
Needs reviewA machine-readable review of install readiness, security metadata, maintenance, and adoption risk.
OpenAgentSkill Trust Score v5
Human review before install
Run only in a sandbox and compare close alternatives before using it for real work.
Stars
92 GitHub stars
Repo activity
92 stars, 5 forks
Maintenance
1d since push
License
MIT
Install
npx skills add xiincs/claude-code-vision-skill
Install safety
standard package or runtime install path
Permission surface
secrets or environment access, shell or command execution
Agent outcomes
No agent outcome data yet
Docs
Strong README/SKILL.md context
Risk summary
Review before production
- Quality score needs review
- Permission surface needs review: secrets or environment access, shell or command execution
- GitHub adoption: 92 GitHub stars
- Stars/forks activity: 92 stars, 5 forks; issue activity unavailable in current metadata
Install readiness
Install path available
- Install path is available
- Repository evidence is available
- License is declared
- No Agent Proven outcome evidence yet
Agent-readable metadata
Machine-readable decision data for this skill.
Use this block or the embedded JSON to decide whether an agent should install this skill, choose an alternative, or ask for human review first.
Suited tasks
- Coding agents workflows
- Claude Code teams
- builders willing to evaluate younger projects
- Inspect source files
Suited agents
Install decision
- Command
- npx skills add xiincs/claude-code-vision-skill
- Policy
- block
- Human review
- yes
Trust and risk
- Trust
- 64/100
- Audit
- 80/100
- Risk level
- Needs review
Outcome loop
- Endpoint
- /api/agent/outcome
- Event ID
- resolve
- Outcomes
- 5
Install command
npx skills add xiincs/claude-code-vision-skillDo not use when
- teams that need a vendor-supported SLA
- high-compliance environments without internal security review
- No OpenAgentSkill engagement data yet
- High-risk permission hints: Shell or command execution, Secrets or environment access
- Dependency or permission surface needs review
Alternative
Grill With Docs
164.7K Stars
npx skills add mattpocock/skills --skill grill-with-docs
Alternative
Code Review
168.6K Stars
npx skills add mattpocock/skills --skill code-review
Alternative
To Spec
164.7K Stars
npx skills add mattpocock/skills --skill to-spec
Alternative
To Tickets
176.7K Stars
npx skills add mattpocock/skills --skill to-tickets
Agent safety v2
40/100 · Avoid automatic install
This skill should not be selected by an agent without explicit human security review.
Do not auto-install. Inspect the source, dependencies, and permission surface first.
high
Shell or command execution
Skill metadata references terminal, CLI, shell, subprocess, or command execution workflows.
medium
Browser automation
Skill may drive a browser or interact with web pages.
medium
Network access
Skill likely fetches remote pages, APIs, repositories, or external services.
high
Secrets or environment access
Skill metadata references credentials, tokens, environment variables, or secret-bearing workflows.
- High-risk permission hints: Shell or command execution, Secrets or environment access
- Dependency or permission surface needs review
Install targets
Install this skill in your agent workflow
Use the public install endpoint to fetch the command, safety checklist, target prompts, and canonical links for this skill.
OpenAgentSkill CLI
Use the registry command when your workflow supports the OpenAgentSkill installer.
$ npx skills add xiincs/claude-code-vision-skillAgent resolve plan
Let an agent verify fit before installing.
The Resolve API returns the selected skill, alternatives, safety policy, audit notes, install target, and copy-paste prompt an agent can follow without scraping this page.
Open JSON
/api/agent/resolve?task=Use%20Claude%20Code%20Vision%20Skill%20for%20an%20agent%20workflow&agent=codex&max_risk=medium
Resolve text
/api/agent/resolve?task=Use%20Claude%20Code%20Vision%20Skill%20for%20an%20agent%20workflow&agent=codex&max_risk=medium&format=text
Install handoff
/api/skills/xiincs-claude-code-vision-skill/install
Agent should check
- Task fit and alternatives from Resolve API.
- Audit score, trust score, and safety policy warnings.
- Install target compatibility for Codex, Claude Code, Cursor, or CLI.
Copy prompt
Task: Use Claude Code Vision Skill in this workspace.
Resolve first: https://www.openagentskill.com/api/agent/resolve?task=Use%20Claude%20Code%20Vision%20Skill%20for%20an%20agent%20workflow&agent=codex&max_risk=medium
Review install handoff: https://www.openagentskill.com/api/skills/xiincs-claude-code-vision-skill/install
Install command: npx skills add xiincs/claude-code-vision-skill
Before running it, summarize audit warnings, required permissions, and the fallback skill if install is risky.Agent handoff
Give an agent the install path, not another directory page.
Use the public install endpoint to fetch the command, safety checklist, target prompts, and canonical links for this skill.
Install handoff
/api/skills/xiincs-claude-code-vision-skill/install
LLM text format
/api/skills/xiincs-claude-code-vision-skill/install?format=text
Find alternatives
/api/skills/search?q=Claude%20Code%20Vision%20Skill&limit=3
Agent prompt
Use Claude Code Vision Skill for this task. Review https://www.openagentskill.com/api/skills/xiincs-claude-code-vision-skill/install, then install with: npx skills add xiincs/claude-code-vision-skillRegistry metadata
Agent-readable profile for automatic skill selection.
This page exposes the same decision, trust, audit, use-case, and install signals through the Registry API, so agents can rank this skill without scraping the UI.
Manifest
/api/registry/manifest/xiincs-claude-code-vision-skill
LLM text
/api/registry/manifest/xiincs-claude-code-vision-skill?format=text
Install alias
/api/registry/install/xiincs-claude-code-vision-skill
Recommend
/api/registry/recommend?task=Use%20Claude%20Code%20Vision%20Skill%20in%20an%20agent%20workflow&limit=3
Agent fit
Coding agents
Use-case tags
Platforms
Python, Claude Code, OpenAI Agents, Browser agents
Audit report
Needs review · 80/100
A machine-readable review of install readiness, security metadata, maintenance, and adoption risk.
Agent decision cockpit
Companion skill for Coding agents
Shortlist this skill and compare it with close alternatives before production adoption.
Role in stack
Companion skill
Primary fit
Coding agents
Trust label
Strong shortlist
Install path
Command ready
Use when
- Coding agents workflows
- Claude Code teams
- builders willing to evaluate younger projects
Evidence
- recent repository activity
- install command or GitHub repo available
- 77/100 quality profile
review first
- No OpenAgentSkill engagement data yet
Implementation path
- 1Install it in a sandbox agent and run one Coding agents task end to end.
- 2Compare output quality, latency, and failure behavior against at least one alternative.
- 3Promote it into production only after reviewing repository permissions, license, and maintenance signals.
Trust profile
Sandbox only
Useful candidate with missing or mixed trust signals. Keep it in an isolated workspace until the outcome loop proves task fit.
GitHub adoption
CHECK92 GitHub stars
Stars/forks activity
CHECK92 stars, 5 forks; issue activity unavailable in current metadata
Recent maintenance
PASS1d since push
License clarity
PASSMIT
Good signals
- AI review approved
- Install path is available
- Repository evidence is available
- Recently maintained repository
- Install command has no obvious high-risk pattern
- Outcome loop is ready but needs first real agent run
Review before install
- Quality score needs review
- Permission surface needs review: secrets or environment access, shell or command execution
- GitHub adoption: 92 GitHub stars
- Stars/forks activity: 92 stars, 5 forks; issue activity unavailable in current metadata
- Dependency/runtime risk: command execution surface, credential or environment access
- Permission surface: secrets or environment access, shell or command execution
- No real agent outcome reports yet
- Human review required before unattended installation
Recommended action
Run only in a sandbox and compare close alternatives before using it for real work.
Quality profile
Strong candidate for agent workflows
Solid option that is likely worth shortlisting for production workflows.
Workflow fit
Use this skill in these scenarios
Build and ship code
Coding agents
I need a coding agent that can understand a repository, edit code, and review pull requests.
Operate web apps
Browser automation
I need my agent to control a browser, fill forms, and verify web app workflows.
Verify behavior
Testing and QA
I need my agent to test a web app, reproduce bugs, and verify fixes.
Workflow fit
Add it to a complete workflow
Operate and verify web apps
Browser QA agent
A workflow for agents that navigate products, fill forms, take screenshots, and verify real user flows across web applications.
Inspect, patch, and verify code
Coding review agent
A workflow for software agents that inspect repositories, review pull requests, generate tests, and turn findings into shippable patches.
Design, build, test, and ship interfaces
Frontend and UI
A practical workflow for agents that turn product briefs or Figma designs into polished frontend code, review the result, test it in a browser, and prepare a safe deployment.
Alternative shortlist
Compare before you install
Similar skills that may fit this task.
Grill With Docs
A relentless interview that pressure-tests a plan against the codebase, sharpens domain language, and updates CONTEXT.md and ADRs when decisions become durable.
Code Review
Review a branch or diff against repository standards and the originating spec in two independent analysis passes.
To Spec
Turn the current conversation and codebase context into a structured implementation spec, then publish it to the configured project issue tracker.
To Tickets
Break a plan, spec, or conversation into independently actionable tracer-bullet tickets with explicit blocking relationships.
Overview
# Claude Code Vision Skill
为 Claude Code 提供多模态视觉能力,支持多种视觉模型分析截图、UI、图表。
专为使用 DeepSeek 等无多模态能力的模型作为 Claude Code 底座的用户设计。
## 截图直贴,所见即所析 遇到 UI 报错、设计稿或数据图表?**直接 alt + v 截图粘贴到 Claude Code 对话中,自动调用视觉模型分析**。无需记住任何命令,像聊天一样自然。
## 串联工具链,自动 UI 审查 配合 **browser-harness** 等工具,可实现**自动截图 → AI 视觉分析 → 生成审查报告 → 修正页面**的完整闭环。前端页面渲染后自动截图,由视觉模型检查布局、样式和交互问题。
## 支持的模型
| Provider | 模型 | API Key 环境变量 | |----------|------|-----------------| | 豆包 (Doubao) | doubao-seed-2-0-pro-260215 | `DOUBAO_API_KEY` | | 通义千问 (Qwen) | qwen-vl-max | `DASHSCOPE_API_KEY` | | OpenAI | gpt-4o | `OPENAI_API_KEY` | | Claude (Anthropic) | claude-sonnet-5 | `ANTHROPIC_API_KEY` | | **任意自定义 provider** | 任意 | `{NAME}_API_KEY` |
不在上面四个内置厂商里的 `--provider` 名字,会按同一套命名约定动态解析:设置 `{NAME}_API_KEY` / `{NAME}_BASE_URL` / `{NAME}_MODEL`(`{NAME}_PROTOCOL` 可选, 默认 `openai`,需要 Anthropic Messages API 形状时设为 `anthropic`),无需改一行代码即可接入任意 OpenAI 兼容端点(vLLM、Ollama、LiteLLM、OpenRouter、Azure OpenAI、自建代理等)。详见 [vision/SKILL.md](vision/SKILL.md#any-custom-provider)。
## 安装依赖
需要 Python 3.10+。
```bash pip install -r requirements.txt ```
`anthropic` 包仅在使用 `--provider anthropic` 时才会被导入,其他 provider 不受影响。
---
## Claude Code 安装流程(给 Claude Code 看的指令)
> **When the user asks you to install this skill, follow these steps IN ORDER. Do NOT skip any step.**
### Step 1 — 询问用户
使用 AskUserQuestion 询问以下信息:
1. **选择 provider**:doubao / qwen / openai / anthropic / 自定义(可多选;自定义需额外要 base URL、model,可选 protocol) 2. **API Key**:每个 provider 的 API key 3. **默认 provider**(多选时):选哪个作为默认
### Step 2 — 运行安装脚本
```bash python install.py \ --api-key qwen:sk-xxx \ --default-provider qwen \ --merge-claude ```
`--api-key` 可重复,`--merge-claude` **必须带上**。
如果用户选了自定义 provider,额外带上 `--base-url` / `--model`(`--protocol` 可选,默认 openai):
```bash python install.py \ --api-key myapi:sk-xxx \ --base-url myapi:https://host/v1 \ --model myapi:my-vision-model \ --default-provider myapi \ --merge-claude ```
### Step 3 — 合并 CLAUDE.md(如果 install.py 未自动完成)
如果未使用 `--merge-cla
Platform compatibility
Technical details
- Version
- 1.0.0
- License
- MIT
- Last updated
- Jul 31, 2026
- Published
- Jul 31, 2026
Frameworks & tools
Decision snapshot
Companion skill
recent repository activity
Audit
Install review
Install and adoption review
- Security
- 77/100
- Maintenance
- 100/100
- Install
- 92/100
Agent-proven evidence
Agent-proven evidence
Outcome reports after resolve, review, install, and one narrow run.
- Success rate
- —
- Recent failure
- —
- Outcomes
- 0
- Output quality
- —
- Failed
- 0
- Not relevant
- 0
- Installs
- 0
- Risk blocked
- 0
- Setup needed
- 0
- Production
- 0
No agent outcome data yet. The first agent run can report success, setup needs, risk blocks, failure, or not-relevant through /api/agent/outcome.
Install
Add to agent workflow
Free and open source. Review the report before installing into production agents.
Growth loop
Share kit
Scenario-led draft for Claude Code Vision Skill, ready for a manual X post.
Claude Code Vision Skill: A reusable skill for Claude Code that adds multimodal vision capabilities for analyzing scree... 92 stars https://www.openagentskill.com/skills/xiincs-claude-code-vision-skill?ref=x
Optional reply with install command
Listing + install path for Claude Code Vision Skill: https://www.openagentskill.com/skills/xiincs-claude-code-vision-skill?ref=x Install: npx skills add xiincs/claude-code-vision-skill
Listing source
Community indexed
This listing was indexed from public sources and is not marked official until a maintainer claim is approved.
- Creator
- xiincs
- Indexed by
- OpenAgentSkill community index
Attribution links to the public repository or creator profile. Creators can claim the listing to update ownership signals.
Claim this skillOwner claim
Claim this skill listing
This Community indexed listing is attributed to xiincs but is not marked official yet. Claim it to add a verified owner signal and make future launch, install, and audit updates easier to trust.
Creator backlink kit
Add the evidence badges to your README
Show the canonical listing, current trust and audit signals, and real Agent-Proven evidence where developers evaluate the repository.
[](https://www.openagentskill.com/skills/xiincs-claude-code-vision-skill)
[](https://www.openagentskill.com/skills/xiincs-claude-code-vision-skill)
[](https://www.openagentskill.com/skills/xiincs-claude-code-vision-skill/audit)
[](https://www.openagentskill.com/skills/xiincs-claude-code-vision-skill)Author
xiincs
@xiincs
Platform fit
Health signals
- GitHub stars
- 92
- Quality score
- 47/100
- Last GitHub push
- Jul 31, 2026
- Framework hints
- 1
- OpenAgentSkill views
- 0
- Install copies
- 0
- Outbound clicks
- 0
Community signal
Share whether this skill looks useful for your agent workflow. Aggregated feedback improves rankings over time.
Trust & safety
Sandbox only
- GitHub adoption92 GitHub starsCHECK
- Stars/forks activity92 stars, 5 forks; issue activity unavailable in current metadataCHECK
- Recent maintenance1d since pushPASS
- License clarityMITPASS
- README/SKILL.md completenessMetadata includes enough usage and workflow contextPASS
- Dependency/runtime riskcommand execution surface, credential or environment accessFIX
Related skills
Grill With Docs
A relentless interview that pressure-tests a plan against the codebase, sharpens domain language, and updates CONTEXT.md and ADRs when decisions become durable.
164.7K Stars · 0 InstallsCode Review
Review a branch or diff against repository standards and the originating spec in two independent analysis passes.
168.6K Stars · 0 InstallsTo Spec
Turn the current conversation and codebase context into a structured implementation spec, then publish it to the configured project issue tracker.
164.7K Stars · 0 InstallsTo Tickets
Break a plan, spec, or conversation into independently actionable tracer-bullet tickets with explicit blocking relationships.
176.7K Stars · 0 Installs