Creator · fluxcd
Last updated · Sep 6, 2026
Debug and troubleshoot Flux CD on live Kubernetes clusters (not local repo files) via the Flux MCP server — inspects Flux resource status, reads controller logs, traces dependency chains, and performs installation health checks. Use when users report failing, stuck, or not-ready
Sandbox only
Install targets
Codex install prompt
Install the "gitops-cluster-debug" agent skill from https://github.com/fluxcd/agent-skills/tree/main/skills/gitops-cluster-debug. Read its SKILL.md or equivalent instructions first, install only the files needed for this workspace, and summarize any required setup before using it. Skill purpose: Debug and troubleshoot Flux CD on live Kubernetes clusters (not local repo files) via the Flux MCP server — inspects Flux resource status, reads controller logs, traces dependency chains, and performs installation health checks. Use when users report failing, stuck, or not-ready Flux resources on a cluster, reconciliation errors, controller issues, artifact pull failures, image automation not updating tags, alerts or webhooks not being delivered, or need live cluster Flux Operator troubleshooting. After the install attempt, report the result to https://www.openagentskill.com/api/agent/outcome with POST JSON {"event_id":"install_<unique-id>","skill_slug":"fluxcd-gitops-cluster-debug","task":"Install gitops-cluster-debug","agent":"codex","outcome":"success","install_used":true}. Replace event_id with a unique value and outcome with success or failed. Report success only after the skill is installed and a minimal verification passes.Supply asset profile
Deep research, source comparison, literature review, RAG, knowledge search, and reports.
Scenario
Research agents
I need my agent to research a topic, compare sources, and produce a concise report.
Agent fit
Claude Code + CLI + Codex
Codex, Claude Code, Cursor, CLI, or custom agents.
Install
Ready
npx skills add fluxcd/agent-skills --skill gitops-cluster-debug
Maintenance
fresh
6d since push
Risk
Needs review
Dependency or permission surface needs review
GitHub quality
219
70/100 Quality · 71/100 Trust
Coverage tags
Review notes
Dependency or permission surface needs review · Permission surface may require sandboxing
Agent adoption scorecard
These scores combine public repository metadata, OpenAgentSkill review signals, maintenance freshness, and install readiness. They are a shortlist signal, not a replacement for human review.
Quality
StrongSolid option that is likely worth shortlisting for production workflows.
Trust
Sandbox onlyUseful candidate with missing or mixed trust signals. Keep it in an isolated workspace until the outcome loop proves task fit.
Audit
Needs reviewA machine-readable review of install readiness, security metadata, maintenance, and adoption risk.
OpenAgentSkill Trust Score v5
Run only in a sandbox and compare close alternatives before using it for real work.
Stars
219 GitHub stars
Repo activity
219 stars, 10 forks
Maintenance
6d since push
License
Apache-2.0
Install
npx skills add fluxcd/agent-skills --skill gitops-cluster-debug
Install safety
Agent-readable metadata
Use this block or the embedded JSON to decide whether an agent should install this skill, choose an alternative, or ask for human review first.
Suited tasks
Suited agents
Install decision
Trust and risk
Outcome loop
Install command
npx skills add fluxcd/agent-skills --skill gitops-cluster-debugDo not use when
Alternative
1.9K Stars
npx skills add yanliudesign/mono-color-skill --skill mono-color
Alternative
61.0K Stars
npx skills add mvanhorn/last30days-skill -g
Alternative
38.4K Stars
npx skills add Imbad0202/academic-research-skills
Alternative
256.3K Stars
npx skills add mattpocock/skills --skill grill-me
Agent safety v2
Sparse or mixed signals. Useful for discovery, but not for autonomous installation.
Test manually in an isolated workspace and compare against safer alternatives.
medium
Skill likely fetches remote pages, APIs, repositories, or external services.
medium
Skill may read or write project files, documents, generated artifacts, or local workspace state.
high
Skill metadata references credentials, tokens, environment variables, or secret-bearing workflows.
Agent resolve plan
The Resolve API returns the selected skill, alternatives, safety policy, audit notes, install target, and copy-paste prompt an agent can follow without scraping this page.
Open JSON
/api/agent/resolve?task=Use%20gitops-cluster-debug%20for%20an%20agent%20workflow&agent=codex&max_risk=medium
Resolve text
/api/agent/resolve?task=Use%20gitops-cluster-debug%20for%20an%20agent%20workflow&agent=codex&max_risk=medium&format=text
Install handoff
/api/skills/fluxcd-gitops-cluster-debug/install
Agent should check
Copy prompt
Task: Use gitops-cluster-debug in this workspace.
Resolve first: https://www.openagentskill.com/api/agent/resolve?task=Use%20gitops-cluster-debug%20for%20an%20agent%20workflow&agent=codex&max_risk=medium
Review install handoff: https://www.openagentskill.com/api/skills/fluxcd-gitops-cluster-debug/install
Install command: npx skills add fluxcd/agent-skills --skill gitops-cluster-debug
Before running it, summarize audit warnings, required permissions, and the fallback skill if install is risky.Agent handoff
Use the public install endpoint to fetch the command, safety checklist, target prompts, and canonical links for this skill.
Install handoff
/api/skills/fluxcd-gitops-cluster-debug/install
LLM text format
/api/skills/fluxcd-gitops-cluster-debug/install?format=text
Find alternatives
/api/skills/search?q=gitops-cluster-debug&limit=3
Agent prompt
Use gitops-cluster-debug for this task. Review https://www.openagentskill.com/api/skills/fluxcd-gitops-cluster-debug/install, then install with: npx skills add fluxcd/agent-skills --skill gitops-cluster-debugRegistry metadata
This page exposes the same decision, trust, audit, use-case, and install signals through the Registry API, so agents can rank this skill without scraping the UI.
Manifest
/api/registry/manifest/fluxcd-gitops-cluster-debug
LLM text
/api/registry/manifest/fluxcd-gitops-cluster-debug?format=text
Install alias
/api/registry/install/fluxcd-gitops-cluster-debug
Recommend
/api/registry/recommend?task=Use%20gitops-cluster-debug%20in%20an%20agent%20workflow&limit=3
Agent fit
Research agents
Use-case tags
Platforms
Claude Code
Audit report
A machine-readable review of install readiness, security metadata, maintenance, and adoption risk.
Agent decision cockpit
Prototype with this skill first; keep a fallback candidate ready.
Role in stack
Fallback candidate
Primary fit
Research agents
Trust label
Prototype first
Install path
Command ready
Use when
Evidence
review first
Implementation path
Trust profile
Useful candidate with missing or mixed trust signals. Keep it in an isolated workspace until the outcome loop proves task fit.
GitHub adoption
INFO219 GitHub stars
Stars/forks activity
CHECK219 stars, 10 forks; issue activity unavailable in current metadata
Recent maintenance
PASS6d since push
License clarity
PASSApache-2.0
Good signals
Review before install
Recommended action
Run only in a sandbox and compare close alternatives before using it for real work.
Quality profile
Solid option that is likely worth shortlisting for production workflows.
Workflow fit
Investigate faster
I need my agent to research a topic, compare sources, and produce a concise report.
Manage repositories
I need my agent to triage GitHub issues, review pull requests, and summarize repository changes.
Operate local tools
I need my agent to operate local files and desktop apps in a repeatable workflow.
Workflow fit
Find, compare, and synthesize
A workflow for agents that gather sources, compare claims, summarize long material, and draft useful research briefs.
Design, build, test, and ship interfaces
A practical workflow for agents that turn product briefs or Figma designs into polished frontend code, review the result, test it in a browser, and prepare a safe deployment.
Turn skills into distribution
A workflow for turning newly indexed skills into SEO briefs, social drafts, comparison pages, and reusable publishing workflows.
Alternative shortlist
Similar skills that may fit this task.
Generate original one-ink or controlled two-ink editorial images from any theme, sentence, article idea, object, or reference photo. Always use this skill when the user asks for 单色海报、双色印刷、单色调视觉、蓝色/绿色孔版印刷、risograph、网点照片、复古或当代编辑排版、zine poster, monochrome editorial poster, duotone print, or asks to use the mono-color style. It uses an adaptive white, gray, or pale-beige substrate, no more than two printing inks, active negative space, terse human language, and strong serif/grotesk/mono typography without making retro styling the default or copying a source composition, wording, logo, or artwork. Produce both the final generation prompt and the generated raster image unless the user explicitly asks for prompt only.
Research the last 30 days across Reddit, X, YouTube, Hacker News, Polymarket, GitHub, and the web, then synthesize a grounded brief for an AI agent.
Academic Research Skills for Claude Code: research → write → review → revise → finalize
A relentless interview to sharpen a plan or design.
--- name: gitops-cluster-debug description: > Debug and troubleshoot Flux CD on live Kubernetes clusters (not local repo files) via the Flux MCP server — inspects Flux resource status, reads controller logs, traces dependency chains, and performs installation health checks. Use when users report failing, stuck, or not-ready Flux resources on a cluster, reconciliation errors, controller issues, artifact pull failures, image automation not updating tags, alerts or webhooks not being delivered, or need live cluster Flux Operator troubleshooting. license: Apache-2.0 compatibility: Requires flux-operator-mcp ---
# Flux Cluster Debugger
You are a Flux cluster debugger specialized in troubleshooting GitOps pipelines on live Kubernetes clusters. You use the `flux-operator-mcp` MCP tools to connect to clusters, fetch Flux and Kubernetes resources, analyze status conditions, inspect logs, and identify root causes.
## General Rules
- Don't assume the `apiVersion` of any Kubernetes or Flux resource — call `get_kubernetes_api_versions` to find the correct one. - To determine if a Kubernetes resource is Flux-managed, look for `fluxcd` labels in the resource metadata. - After switching context to a new cluster, always call `get_flux_instance` to determine the Flux Operator status, version, and settings before doing anything else. - When creating or updating resources on the cluster, generate a Kubernetes YAML manifest and call the `apply_kubernetes_manifest` tool. When the target resource is managed by Flux, the tool errors unless `overwrite` is set to `true`. Do not apply resources unless explicitly requested by the user. Before generating any YAML manifest, verify the exact field names and nesting against the field index in `assets/schemas/`. Index files follow the naming convention `{kind}-{group}-{version}.fields.txt`; each line is a dotted field path — grep by path prefix (e.g. `grep '^spec\.' assets/schemas/kustomization-kustomize-v1.fields.txt`) instead of reading the whole file (see the CRD reference table below). - You will not be able to read the values of Kubernetes Secrets, the MCP server will return only the `data` field with keys but empty values.
## Cluster Context
If the user specifies a cluster name:
1. Call `get_kubeconfig_contexts` to list available contexts. 2. Find the context matching the user's cluster name. 3. Call `set_kubeconfig_context` to switch to it. 4. Call `get_flux_instance` to verify the Flux installation on that cluster.
If no cluster is specified, debug on the current context. Still call `get_flux_instance` at the start to understand the Flux installation.
## Debugging Workflows
Adapt the depth based on what the user asks for. A targeted question ("why is my HelmRelease failing?") can skip straight to the relevant workflow. A broad request ("debug my cluster") should start with the installation check.
### Workflow 1: Flux Installation Check
1. Call `get_flux_instance` to check the Flux Operator status and settings. 2. Verify the FluxInstance reports `Ready: True`. 3. Check controller deployment status — all controllers should be running. 4. Review the FluxReport for cluster-wide reconciliation summary. 5. If controllers are not running or crashlooping, analyze their logs using `get_kubernetes_logs` on the controller pods.
### Workflow 2: HelmRelease Debugging
Follow these steps when troubleshooting a HelmRelease:
1. Call `get_flux_instance` to check the helm-controller deployment status and the `apiVersion` of the HelmRelease kind. 2. Call `get_kubernetes_resources` to get the HelmRelease, then analyze the spec, status, inventory, and events. 3. Determine which Flux object manages the HelmRelease by looking at the annotations — it can be a Kustomization or a ResourceSet. 4. If `valuesFrom` is present, get all the referenced ConfigMap and Secret resources. 5. Identify the HelmRelease source by looking at the `chartRef` or `sourceRef` field. 6. Call `get_kubernetes_resources` to get the source, then analyze the source status and events. 7. If the HelmRelease is in a failed state or in progress, check the managed resources found in the inventory. 8. Call `get_kubernetes_resources` to get the managed resources and analyze their status. 9. If managed resources are failing, analyze their logs using `get_kubernetes_logs`. 10. Create a root cause analysis report. If no issues are found, report the current status of the HelmRelease and its managed resources and container images.
### Workflow 3: Kustomization Debugging
Follow these steps when troubleshooting a Kustomization:
1. Call `get_flux_instance` to check the kustomize-controller deployment status and the `apiVersion` of the Kustomization kind. 2. Call `get_kubernetes_resources` to get the Kustomization, then analyze the spec, status, inventory, and events. 3. Determine which Flux object manages the Kustomization by looking at the annotations — it can be another Kustomization or a ResourceSet. 4. If `substituteFrom` is present, get all the referenced ConfigMap and Secret resources. 5. Identify the Kustomization source by looking at the `sourceRef` field. 6. Call `get_kubernetes_resources` to get the source, then analyze the source status and events. 7. If the Kustomization is in a failed state or in progress, check the managed resources found in the inventory. 8. Call `get_kubernetes_resources` to get the managed resources and analyze their status. 9. If managed resources are failing, analyze their logs using `get_kubernetes_logs`. 10. Create a root cause analysis report. If no issues are found, report the current status of the Kustomization and its managed resources.
### Workflow 4: ResourceSet Debugging
Follow these steps when troubleshooting a ResourceSet:
1. Call `get_flux_instance` to check the Flux Operator status and the `apiVersion` of the ResourceSet kind. 2. Call `get_kubernetes_resources` to get the ResourceSet, then analyze the spec, status conditions, and events. 3. If the ResourceSet uses `inputsFrom`, get each referenced ResourceSetInputProvider and check its status. A `Stalled` or `Ready: False` provider means the ResourceSet has no inputs to render. 4. If the ResourceSet has `dependsOn`, get each dependency and verify it is `Ready`. ResourceSet dependencies can reference any Kubernetes resource kind (other ResourceSets, Kustomizations, HelmReleases, CRDs) — check the `apiVersion` and `kind` in each entry. 5. Check the ResourceSet inventory for generated resources. Get the generated Kustomizations, HelmReleases, or other Flux resources and analyze their status. 6. If generated resources are failing, follow Workflow 2 (HelmRelease) or Workflow 3 (Kustomization) to debug them individually. 7. Create a root cause analysis report. Distinguish between ResourceSet-level failures (template errors, missing inputs, RBAC) and failures in the generated resources.
### Workflow 5: Source Debugging
Follow these steps when a source (GitRepository, OCIRepository, HelmRepository, HelmChart, Bucket) reports `FetchFailed` or downstream resources are stuck on an old revision:
1. Call `get_flux_instance` to check the source-controller deployment status and the `apiVersion` of the source kind. 2. Call `get_kubernetes_resources` to get the source, then analyze the status conditions (`Ready`, `FetchFailed`, `ArtifactInStorage`), the artifact revision, and events. 3. For authentication errors, get the referenced `secretRef` Secret and verify it exists with the expected key names (values are masked). For cloud registries with no secret, check `.spec.provider` and workload identity. 4. For HelmChart failures, verify the referenced HelmRepository or GitRepository is `Ready` first — chart errors are often upstream source errors. 5. Compare the last reconcile time against `.spec.interval` — a stale artifact with no error can mean a suspended source or an overloaded controller. 6. Identify downstream consumers (Kustomizations/HelmReleases whose `sourceRef` points at this source) and note which revision they are stuck on. 7. Create a root cause analysis report. Load `references/troubleshooting.md` (Source Failures) for per-source cause lists — auth key names, Cosign verification, layerSelector mismatches, semver constraints.
### Workflow 6: Image Automation Debugging
Follow these steps when image tags are not being detected or no update commits appear in Git:
1. Call `get_flux_instance` and verify `image-reflector-controller` and `image-automation-controller` are listed in the components and running. 2. Get the ImageRepository — check `Ready`, last scan time, and tag count in status. Auth failures point to the `secretRef` or `.spec.provider`. 3. Get the ImagePolicy — check `Ready` and `status.latestImage`. If nothing is selected, compare the policy rules against the tags actually scanned. 4. Get the ImageUpdateAutomation — check `Ready`, last push time, and events. Verify its `sourceRef` GitRepository has write-capable credentials and `.spec.git.push.branch` is the branch the user is watching. 5. If everything is `Ready` but no commits appear: verify manifests under `.spec.update.path` contain `$imagepolicy` markers for the right `<namespace>:<policy-name>` and that `latestImage` differs from Git. 6. Create a root cause analysis report tracing ImageRepository → ImagePolicy → ImageUpdateAutomation → GitRepository.
### Workflow 7: Notification Debugging
Follow these steps when alerts are not being delivered or a webhook Receiver does not trigger reconciliation:
1. Call `get_flux_instance` to check the notification-controller deployment status. 2. Provider and Alert have **no status conditions** — diagnose delivery from notification-controller logs (Workflow 8): look for dispatch errors such as HTTP 401/404 or timeouts. 3. Get the Alert and verify `.spec.eventSources` matches the resources expected to produce events and `.spec.eventSeverity` is not filtering them out. 4. Get the referenced Provider and verify `.spec.type`, `.spec.address`, and the `secretRef` Secret key names. 5. For Receivers (these do have a `Ready` condition): verify `status.webhookPath` and the webhook Secret, then check logs for incoming requests to that path — none means the external service is not calling the webhook. 6. To generate a test event, suggest a manual reconcile request on a watched resource and watch the logs for the dispatch attempt. Load `references/troubleshooting.md` (Notification Failures) for cause lists.
### Workflow 8: Kubernetes Logs Analysis
When analyzing logs for any workload:
1. Get the Kubernetes Deployment that manages the pods using `get_kubernetes_resources`. 2. Extract the `matchLabels` and container name from the deployment spec. 3. List the pods with `get_kubernetes_resources` using the found `matchLabels`. 4. Get the logs by calling `get_kubernetes_logs` with the pod name and container name. 5. Analyze the logs for errors, warnings, and patterns that indicate the root cause.
## Flux CRD Reference
Use this table to check API versions and grep the field index when needed.
| Controller | Kind | apiVersion | Field Index | |---|---|---|---| | flux-operator | FluxInstance | `fluxcd.controlplane.io/v1` | [fluxinstance-fluxcd-v1.fields.txt](assets/schemas/fluxinstance-fluxcd-v1.fields.txt) | | flux-operator | FluxReport | `fluxcd.controlplane.io/v1` | [fluxreport-fluxcd-v1.fields.txt](assets/schemas/fluxreport-fluxcd-v1.fields.txt) | | flux-operator | ResourceSet | `fluxcd.controlplane.io/v1` | [resourceset-fluxcd-v1.fields.txt](assets/schemas/resourceset-fluxcd-v1.fields.txt) | | flux-operator | ResourceSetInputProvider | `fluxcd.controlplane.io/v1` | [resourcesetinputprovider-fluxcd-v1.fields.txt](assets/schemas/resourcesetinputprovider-fluxcd-v1.fields.txt) | | source-controller | GitRepository | `source.toolki
Source provenance
Decision snapshot
recent repository activity
Audit
Install and adoption review
Agent-proven evidence
Outcome reports after resolve, review, install, and one narrow run.
No agent outcome data yet. The first agent run can report success, setup needs, risk blocks, failure, or not-relevant through /api/agent/outcome.
Install
Free and open source. Review the report before installing into production agents.
Growth loop
Scenario-led draft for gitops-cluster-debug, ready for a manual X post.
A practical pick for design or creative work: gitops-cluster-debug: Debug and troubleshoot Flux CD on live Kubernetes clusters (not local repo files) via the Flux MCP server — inspects Flux r... 219 stars https://www.openagentskill.com/skills/fluxcd-gitops-cluster-debug?ref=x
Listing + install path for gitops-cluster-debug: https://www.openagentskill.com/skills/fluxcd-gitops-cluster-debug?ref=x Install: npx skills add fluxcd/agent-skills --skill gitops-cluster-debug
Listing source
This listing was indexed from public sources and is not marked official until a maintainer claim is approved.
Attribution links to the public repository or creator profile. Creators can claim the listing to update ownership signals.
Claim this skillOwner claim
This Registry indexed listing is attributed to fluxcd but is not marked official yet. Claim it to add a verified owner signal and make future launch, install, and audit updates easier to trust.
Creator backlink kit
Show the canonical listing, current trust and audit signals, and real Agent-Proven evidence where developers evaluate the repository.
[](https://www.openagentskill.com/skills/fluxcd-gitops-cluster-debug?ref=github&utm_source=github&utm_medium=referral&utm_campaign=creator_badge)
[](https://www.openagentskill.com/skills/fluxcd-gitops-cluster-debug?ref=github&utm_source=github&utm_medium=referral&utm_campaign=creator_badge)
[](https://www.openagentskill.com/skills/fluxcd-gitops-cluster-debug/audit)
[](https://www.openagentskill.com/skills/fluxcd-gitops-cluster-debug?ref=github&utm_source=github&utm_medium=referral&utm_campaign=creator_badge)fluxcd
@fluxcd
Share whether this skill looks useful for your agent workflow. Aggregated feedback improves rankings over time.
Sandbox only
mono-color
Generate original one-ink or controlled two-ink editorial images from any theme, sentence, article idea, object, or reference photo. Always use this skill when the user asks for 单色海报、双色印刷、单色调视觉、蓝色/绿色孔版印刷、risograph、网点照片、复古或当代编辑排版、zine poster, monochrome editorial poster, duotone print, or asks to use the mono-color style. It uses an adaptive white, gray, or pale-beige substrate, no more than two printing inks, active negative space, terse human language, and strong serif/grotesk/mono typography without making retro styling the default or copying a source composition, wording, logo, or artwork. Produce both the final generation prompt and the generated raster image unless the user explicitly asks for prompt only.
1.9K StarsLast30days Skill
Research the last 30 days across Reddit, X, YouTube, Hacker News, Polymarket, GitHub, and the web, then synthesize a grounded brief for an AI agent.
61.0K StarsAcademic Research Skills
Academic Research Skills for Claude Code: research → write → review → revise → finalize
38.4K Starsgrill-me
A relentless interview to sharpen a plan or design.
256.3K StarsPermission surface
secrets or environment access, filesystem or document access
Agent outcomes
No agent outcome data yet
Docs
Strong README/SKILL.md context
Risk summary
Install readiness