Creator · awslabs
Last updated · Sep 4, 2026
This skill should be used when the user asks to \"analyze this codebase\", \"document this service\", \"generate technical docs\", \"I inherited this code\", \"help me understand this system\", \"create docs for this project\", \"what does this system look like\", \"onboard me to
Sandbox only
Install targets
Codex install prompt
Install the "document-service" agent skill from https://github.com/awslabs/agent-plugins/tree/main/plugins/codebase-documentor-for-aws/skills/document-service. Read its SKILL.md or equivalent instructions first, install only the files needed for this workspace, and summarize any required setup before using it. Skill purpose: This skill should be used when the user asks to \"analyze this codebase\", \"document this service\", \"generate technical docs\", \"I inherited this code\", \"help me understand this system\", \"create docs for this project\", \"what does this system look like\", \"onboard me to this codebase\", \"this codebase has no docs\", \"visualize the architecture from code\", or any explicit request to produce structured documentation or architecture diagrams from an existing codebase. Specifically optimized for AWS workloads (CDK, CloudFormation, Terraform) with source-of-truth citations. Do NOT activate for code reviews, single-function explanations, generating new code, or general coding tasks. After the install attempt, report the result to https://www.openagentskill.com/api/agent/outcome with POST JSON {"event_id":"install_<unique-id>","skill_slug":"awslabs-document-service","task":"Install document-service","agent":"codex","outcome":"success","install_used":true}. Replace event_id with a unique value and outcome with success or failed. Report success only after the skill is installed and a minimal verification passes.Supply asset profile
Deep research, source comparison, literature review, RAG, knowledge search, and reports.
Scenario
Document processing
I need my agent to read PDFs, extract tables, and turn documents into structured data.
Agent fit
Claude Code + OpenAI Agents + Cursor
Codex, Claude Code, Cursor, CLI, or custom agents.
Install
Ready
npx skills add awslabs/agent-plugins --skill document-service
Maintenance
fresh
3d since push
Risk
Needs review
Dependency or permission surface needs review
GitHub quality
888
77/100 Quality · 77/100 Trust
Coverage tags
Review notes
Dependency or permission surface needs review · Permission surface may require sandboxing
Agent adoption scorecard
These scores combine public repository metadata, OpenAgentSkill review signals, maintenance freshness, and install readiness. They are a shortlist signal, not a replacement for human review.
Quality
StrongSolid option that is likely worth shortlisting for production workflows.
Trust
Sandbox onlyUseful candidate with missing or mixed trust signals. Keep it in an isolated workspace until the outcome loop proves task fit.
Audit
Needs reviewA machine-readable review of install readiness, security metadata, maintenance, and adoption risk.
OpenAgentSkill Trust Score v5
Run only in a sandbox and compare close alternatives before using it for real work.
Stars
888 GitHub stars
Repo activity
888 stars, 152 forks
Maintenance
3d since push
License
Apache-2.0
Install
npx skills add awslabs/agent-plugins --skill document-service
Install safety
Agent-readable metadata
Use this block or the embedded JSON to decide whether an agent should install this skill, choose an alternative, or ask for human review first.
Suited tasks
Suited agents
Install decision
Trust and risk
Outcome loop
Install command
npx skills add awslabs/agent-plugins --skill document-serviceDo not use when
Alternative
1.9K Stars
npx skills add yanliudesign/mono-color-skill --skill mono-color
Alternative
61.0K Stars
npx skills add mvanhorn/last30days-skill -g
Alternative
38.4K Stars
npx skills add Imbad0202/academic-research-skills
Alternative
256.3K Stars
npx skills add mattpocock/skills --skill grill-me
Agent safety v2
Sparse or mixed signals. Useful for discovery, but not for autonomous installation.
Test manually in an isolated workspace and compare against safer alternatives.
medium
Skill likely fetches remote pages, APIs, repositories, or external services.
medium
Skill may read or write project files, documents, generated artifacts, or local workspace state.
high
Skill metadata references credentials, tokens, environment variables, or secret-bearing workflows.
medium
Skill may inspect schemas, query databases, or work with persistent stores.
Agent resolve plan
The Resolve API returns the selected skill, alternatives, safety policy, audit notes, install target, and copy-paste prompt an agent can follow without scraping this page.
Open JSON
/api/agent/resolve?task=Use%20document-service%20for%20an%20agent%20workflow&agent=codex&max_risk=medium
Resolve text
/api/agent/resolve?task=Use%20document-service%20for%20an%20agent%20workflow&agent=codex&max_risk=medium&format=text
Install handoff
/api/skills/awslabs-document-service/install
Agent should check
Copy prompt
Task: Use document-service in this workspace.
Resolve first: https://www.openagentskill.com/api/agent/resolve?task=Use%20document-service%20for%20an%20agent%20workflow&agent=codex&max_risk=medium
Review install handoff: https://www.openagentskill.com/api/skills/awslabs-document-service/install
Install command: npx skills add awslabs/agent-plugins --skill document-service
Before running it, summarize audit warnings, required permissions, and the fallback skill if install is risky.Agent handoff
Use the public install endpoint to fetch the command, safety checklist, target prompts, and canonical links for this skill.
Install handoff
/api/skills/awslabs-document-service/install
LLM text format
/api/skills/awslabs-document-service/install?format=text
Find alternatives
/api/skills/search?q=document-service&limit=3
Agent prompt
Use document-service for this task. Review https://www.openagentskill.com/api/skills/awslabs-document-service/install, then install with: npx skills add awslabs/agent-plugins --skill document-serviceRegistry metadata
This page exposes the same decision, trust, audit, use-case, and install signals through the Registry API, so agents can rank this skill without scraping the UI.
Manifest
/api/registry/manifest/awslabs-document-service
LLM text
/api/registry/manifest/awslabs-document-service?format=text
Install alias
/api/registry/install/awslabs-document-service
Recommend
/api/registry/recommend?task=Use%20document-service%20in%20an%20agent%20workflow&limit=3
Agent fit
Workflow automation
Use-case tags
Platforms
Claude Code, OpenAI Agents, Cursor
Audit report
A machine-readable review of install readiness, security metadata, maintenance, and adoption risk.
Agent decision cockpit
Use this as a leading candidate, then validate the README and install path in your own agent stack.
Role in stack
Primary pick
Primary fit
Workflow automation
Trust label
Production-ready
Install path
Command ready
Use when
Evidence
review first
Implementation path
Trust profile
Useful candidate with missing or mixed trust signals. Keep it in an isolated workspace until the outcome loop proves task fit.
GitHub adoption
INFO888 GitHub stars
Stars/forks activity
INFO888 stars, 152 forks; issue activity unavailable in current metadata
Recent maintenance
PASS3d since push
License clarity
PASSApache-2.0
Good signals
Review before install
Recommended action
Run only in a sandbox and compare close alternatives before using it for real work.
Quality profile
Solid option that is likely worth shortlisting for production workflows.
Workflow fit
Automate repeated work
I need my agent to automate a repeated workflow across tools and files.
Parse messy files
I need my agent to read PDFs, extract tables, and turn documents into structured data.
Build and ship code
I need a coding agent that can understand a repository, edit code, and review pull requests.
Workflow fit
Turn skills into distribution
A workflow for turning newly indexed skills into SEO briefs, social drafts, comparison pages, and reusable publishing workflows.
Ingest, retrieve, and cite
A workflow for document-heavy agents that ingest files, create searchable knowledge, retrieve relevant context, and answer with grounded sources.
Find, compare, and synthesize
A workflow for agents that gather sources, compare claims, summarize long material, and draft useful research briefs.
Alternative shortlist
Similar skills that may fit this task.
Generate original one-ink or controlled two-ink editorial images from any theme, sentence, article idea, object, or reference photo. Always use this skill when the user asks for 单色海报、双色印刷、单色调视觉、蓝色/绿色孔版印刷、risograph、网点照片、复古或当代编辑排版、zine poster, monochrome editorial poster, duotone print, or asks to use the mono-color style. It uses an adaptive white, gray, or pale-beige substrate, no more than two printing inks, active negative space, terse human language, and strong serif/grotesk/mono typography without making retro styling the default or copying a source composition, wording, logo, or artwork. Produce both the final generation prompt and the generated raster image unless the user explicitly asks for prompt only.
Research the last 30 days across Reddit, X, YouTube, Hacker News, Polymarket, GitHub, and the web, then synthesize a grounded brief for an AI agent.
Academic Research Skills for Claude Code: research → write → review → revise → finalize
A relentless interview to sharpen a plan or design.
--- name: document-service description: "This skill should be used when the user asks to \"analyze this codebase\", \"document this service\", \"generate technical docs\", \"I inherited this code\", \"help me understand this system\", \"create docs for this project\", \"what does this system look like\", \"onboard me to this codebase\", \"this codebase has no docs\", \"visualize the architecture from code\", or any explicit request to produce structured documentation or architecture diagrams from an existing codebase. Specifically optimized for AWS workloads (CDK, CloudFormation, Terraform) with source-of-truth citations. Do NOT activate for code reviews, single-function explanations, generating new code, or general coding tasks." license: Apache-2.0 ---
# Document Service
Analyze codebases to produce structured technical documentation and architecture diagrams with source-of-truth citations. Every finding links back to the exact file and line it was derived from. Optimized for AWS workloads but works with any codebase.
## Core Principles
- **Explain WHY, not just WHAT.** The reader inherited this codebase and has zero context. Listing components is not enough — explain why the architecture is shaped this way. Search for code comments, TODOs, and commit messages that reveal design rationale. When no rationale exists, mark it `[RATIONALE UNKNOWN]`. - **Trace end-to-end flows.** For every API endpoint or message handler, trace the complete request path from entry to response. Note every intermediate step, transformation, timeout, and failure point. This is the "if it breaks at 3am, where do I look?" analysis. - **Deep-dive complex logic.** Identify the most complex or domain-specific code paths (ML pipelines, business rule engines, state machines, custom algorithms). Document HOW they work at the implementation level — the algorithm, key parameters, edge cases, and where production bugs will occur. Surface-level summaries of complex code provide no value over a naive AI prompt. - **Surface implicit knowledge.** Look for hardcoded values, magic numbers, environment-dependent behavior, and undocumented assumptions. These are the tribal knowledge items that disappear when teams leave. - **Every claim must be traceable.** Include `file:line` citations for every finding. See [citation-format.md](references/citation-format.md). Verify citations precisely — re-read the cited file and confirm the line number is within ±3 lines. Anchor with function/variable names. - **Code is the source of truth.** Document what actually exists in code, not what READMEs or wikis claim. Flag every discrepancy between documentation and reality. - **Mark unknowns and risks explicitly.** Use `[UNKNOWN]` for items not inferable from code, `[RISK]` for unhandled failure modes, `[INFERRED]` for educated guesses, `[RATIONALE UNKNOWN]` for unexplained architecture choices. Omitting markers undermines trust. - **Verify quantitative claims.** List directory entries programmatically and use exact counts.
## Workflow
The workflow runs autonomously from Step 2 onward. Step 1 is the only interactive step.
### Step 1: Gather Context
Gather from the user:
- Target directory or service to analyze - Any existing documentation, design docs, or business context (accept "nothing" — this skill is designed for undocumented codebases)
If existing docs are provided, read them first to establish baseline context. If the target directory and context are already known (e.g., provided via automation or a pre-configured prompt), skip the interactive step and proceed directly to Step 2.
Check whether `CODEBASE_ANALYSIS.md` already exists at the output path. If so, ask the user: "Overwrite or write to a different filename?" Resolve this before proceeding — the rest of the workflow runs autonomously.
### Step 2: Build File Tree and Detect Project Type
1. List all files recursively in the target directory 2. Apply exclusion patterns from [exclusion-patterns.md](references/exclusion-patterns.md). Also respect `.gitignore`. 3. Detect project type and framework from characteristic files. See [discovery-patterns.md](references/discovery-patterns.md). 4. Identify entry points based on detected project type. See [discovery-patterns.md](references/discovery-patterns.md). 5. Read the README, CLAUDE.md, or AGENTS.md if present — these contain project context. 6. Check git branch names (`git branch -a`) for strategic context (e.g., a `dev/rust` branch signals a language migration in progress). Note active branches in the Architecture Overview.
### Step 3: Generate Documentation Outline
Produce a hierarchical outline mapping each documentation section to specific source files:
```markdown ## Documentation Outline
1. Architecture Overview → [entry points, IaC stack files] — explain WHY, not just WHAT 2. [Module A: detected name] → [source files for module A] 3. [Module B: detected name] → [source files for module B] 4. Shared Utilities → [shared/common source files] 5. Request Lifecycle → [trace end-to-end flows through the system] 6. Domain Logic Deep-Dive → [core services at implementation level: algorithms, parameters, edge cases] 7. Startup and Initialization → [boot sequence, model loading, cache warmup, dependency checks] 8. API Contracts → [route definitions, OpenAPI specs] 9. Data Models → [schema files, ORM models] 10. Deployment → [IaC files, Dockerfiles] 11. Configuration → [config files, .env.example, prompt templates, YAML configs, secrets refs] 12. Monitoring and Observability → [log groups, metrics, tracing, alarms, dashboards] 13. Security → [auth, encryption, IAM, network isolation] 14. Local Development → [how to run/test locally, CPU fallback, dev environment setup] 15. Discrepancies → (cross-reference README/metadata vs actual code) 16. Failure Modes → (cross-cutting — include detection + recovery) 17. Timeout and Dependency Chain → (map cascading timeouts across layers) ```
Follow the section structure in [technical-doc-template.md](references/technical-doc-template.md) but adapt to the actual codebase — add sections for significant modules, skip sections that don't apply. Aim for balance: each section should map to a meaningful subset of files. If a module maps to more than ~30 files, consider splitting it into sub-sections.
**Do NOT pause for user review.** Proceed immediately to analysis.
### Step 4: Analyze
Two core analysis paths:
#### Path A: Application Code
For each outline section, read mapped source files and extract:
1. API and service definitions — route handlers, controllers, gRPC services, GraphQL resolvers 2. Data model definitions — database schemas, ORM models, type definitions 3. Internal dependencies — imports between modules, shared utilities, event handlers 4. External integrations — SDK clients, HTTP calls, queue producers/consumers 5. Configuration — environment variables, feature flags, secrets references
Consult [framework-patterns.md](references/framework-patterns.md) for framework-specific extraction patterns.
#### Path B: Infrastructure-as-Code
When IaC files are detected (CDK, CloudFormation, Terraform, Serverless Framework):
1. Parse resource definitions. Identify AWS resource types, relationships, and networking topology. 2. Map infrastructure to application components that use them. 3. Extract networking topology — VPCs, subnets, security groups. 4. Consult MCP servers — use `awsiac` to confirm resource interpretations, `awsknowledge` for service descriptions.
When no IaC is found, infer infrastructure from application code (SDK clients, connection strings, environment variables) and mark components as `[INFERRED]`.
**Note on CDK projects:** In CDK codebases, the IaC IS application code (TypeScript/Python constructs). Process CDK files in a single pass covering both Path A and Path B rather than treating them as separate analyses. Extract both the resource definitions (Path B) and the application logic interleaved with them (Lambda bundling, environment wiring, IAM grants — Path A) simultaneously.
#### Writing Sections
For each outline section:
1. Re-read mapped source files for exact line numbers — do not rely on memory from earlier steps. 2. Use grep for patterns — route definitions, model declarations, error handlers. 3. Write content with inline citations. See [citation-format.md](references/citation-format.md). 4. **Document every source file.** Enumerate ALL non-generated source files. Every file should appear somewhere in the documentation — in a module table, component table, or at minimum a file inventory. Files that define symbols never imported by any execution path should be flagged as `[UNUSED]` potential dead code. 5. **Analyze the test suite.** Document what tests verify, what coverage gaps exist, and how to interpret test failures. Tests reveal expected behavior and edge cases.
Process cross-cutting sections (Failure Modes, Configuration, Security, Discrepancies) last, drawing on accumulated knowledge.
**Discrepancy detection**: After analyzing the codebase, re-read the README, CLAUDE.md, package.json description, and any project metadata. Flag every claim that does not match the actual code — features referenced but not implemented, resource types that differ, architecture components that don't exist. For legacy codebases, this "trust but verify" pass is the single most valuable output.
**Actionable failure modes**: For each failure mode, include the detection method (CloudWatch metric, log pattern, symptom) and recovery steps (actual commands), not just a description. The reader is an on-call engineer at 3am.
#### Deep Analysis Approach
Do not attempt a single-pass skim. For each module or service, use iterative deepening:
1. **First pass** — scan file structure and entry points to understand scope 2. **Second pass** — read core files, identify questions (what calls this? where is this configured? what happens on error?) 3. **Third pass** — search for answers to those questions across the codebase, trace cross-module dependencies 4. **Write** — only write the section after all three passes. Re-read cited files to verify exact line numbers.
#### Large Codebase Strategy
For codebases with multiple top-level modules, deep nesting, or hundreds of source files:
- **Primary: tracked sequential analysis.** Create a `.codebase-documentor-progress.md` task board to track progress through sections, enabling resumability if interrupted. This works on all platforms (Claude Code, Cursor, Codex, or any coding assistant). - **Acceleration: parallel workers.** If the environment supports spawning parallel agents, assign outline sections to independent workers. Each worker reads its mapped files and produces section content with citations. Keep Architecture Overview and cross-cutting sections in the main session for assembly.
See [recursive-analysis.md](references/recursive-analysis.md) for detailed instructions on both approaches.
### Step 5: Generate Diagrams
Two types of diagrams serve different purposes:
**Sequence/flow diagrams — inline Mermaid.** For request lifecycle traces and data pipeline flows identified in Step 4, generate Mermaid `sequenceDiagram` or `flowchart` blocks inline in the relevant CODEBASE_ANALYSIS.md sections. Mermaid is the community standard for simple flow diagrams and renders natively on GitHub. Keep these focused — one diagram per major request path or data flow.
**Architecture diagram — always attempt the `aws-architecture-diagram` skill first.** For the system-level architecture diagram (services, infrastructure, boundaries): invoke the `aws-architecture-diagram` skill (part of the `deploy-on-aws` plugin) with "analyze [target-directory]" to trigger Mode A. It produces a validated draw.io diagram (`docs/*.drawio`) with official AWS4 icons and professional styling. **Only if** the skill is genuinely unavailable (not installed, invocation fails), fall back to a Mermaid `flowchart TD` architecture overview directly
Source provenance
Decision snapshot
888 GitHub stars
Audit
Install and adoption review
Agent-proven evidence
Outcome reports after resolve, review, install, and one narrow run.
No agent outcome data yet. The first agent run can report success, setup needs, risk blocks, failure, or not-relevant through /api/agent/outcome.
Install
Free and open source. Review the report before installing into production agents.
Growth loop
Scenario-led draft for document-service, ready for a manual X post.
document-service: This skill should be used when the user asks to \"analyze this codebase\", \"document this se... 888 stars https://www.openagentskill.com/skills/awslabs-document-service?ref=x
Listing + install path for document-service: https://www.openagentskill.com/skills/awslabs-document-service?ref=x Install: npx skills add awslabs/agent-plugins --skill document-service
Listing source
This listing was indexed from public sources and is not marked official until a maintainer claim is approved.
Attribution links to the public repository or creator profile. Creators can claim the listing to update ownership signals.
Claim this skillOwner claim
This Registry indexed listing is attributed to awslabs but is not marked official yet. Claim it to add a verified owner signal and make future launch, install, and audit updates easier to trust.
Creator backlink kit
Show the canonical listing, current trust and audit signals, and real Agent-Proven evidence where developers evaluate the repository.
[](https://www.openagentskill.com/skills/awslabs-document-service?ref=github&utm_source=github&utm_medium=referral&utm_campaign=creator_badge)
[](https://www.openagentskill.com/skills/awslabs-document-service?ref=github&utm_source=github&utm_medium=referral&utm_campaign=creator_badge)
[](https://www.openagentskill.com/skills/awslabs-document-service/audit)
[](https://www.openagentskill.com/skills/awslabs-document-service?ref=github&utm_source=github&utm_medium=referral&utm_campaign=creator_badge)awslabs
@awslabs
Share whether this skill looks useful for your agent workflow. Aggregated feedback improves rankings over time.
Sandbox only
mono-color
Generate original one-ink or controlled two-ink editorial images from any theme, sentence, article idea, object, or reference photo. Always use this skill when the user asks for 单色海报、双色印刷、单色调视觉、蓝色/绿色孔版印刷、risograph、网点照片、复古或当代编辑排版、zine poster, monochrome editorial poster, duotone print, or asks to use the mono-color style. It uses an adaptive white, gray, or pale-beige substrate, no more than two printing inks, active negative space, terse human language, and strong serif/grotesk/mono typography without making retro styling the default or copying a source composition, wording, logo, or artwork. Produce both the final generation prompt and the generated raster image unless the user explicitly asks for prompt only.
1.9K StarsLast30days Skill
Research the last 30 days across Reddit, X, YouTube, Hacker News, Polymarket, GitHub, and the web, then synthesize a grounded brief for an AI agent.
61.0K StarsAcademic Research Skills
Academic Research Skills for Claude Code: research → write → review → revise → finalize
38.4K Starsgrill-me
A relentless interview to sharpen a plan or design.
256.3K StarsPermission surface
secrets or environment access, filesystem or document access
Agent outcomes
No agent outcome data yet
Docs
Strong README/SKILL.md context
Risk summary
Install readiness