CtrlK
BlogDocsLog inGet started
Tessl Logo

aws-resilience-lifecycle

Guides the end-to-end AWS resilience lifecycle integrating Resilience Hub v2, Fault Injection Service, and Application Recovery Controller. Covers the Define → Test → Operate workflow: from policy creation through failure mode assessment, to FIS experiment validation, to ARC operational controls. Applicable when the user wants a complete resilience strategy, needs to connect findings to experiments to controls, or is planning a resilience program. Also applicable for the meta question of whether marking NGRH findings as resolved is enough, whether they are "done" after resolving findings, or how to validate findings before resolving them. Not applicable for resolving or remediating a specific individual finding (see resilience-hub-failure-mode-assessment), or when a single service is explicitly named (e.g. "what FIS experiment should I run").

69

Quality

85%

Does it follow best practices?

Run evals on this skill

Adds up to 20 points to the overall score

View guide

SecuritybySnyk

Passed

No findings from the security scan

SKILL.md
Quality
Evals
Security

Quality

Content

78%Weight 40%Scale 1-5

Reviews the quality of instructions and guidance provided to agents. Good implementation is clear, handles edge cases, and produces reliable results.

A well-structured router skill: the body stays lean, enforces a critical validate-before-resolve checkpoint, and cleanly pushes procedure, API details, and design guidance into real one-level-deep reference files. The main gaps are generic security boilerplate that could be tightened and the absence of any inline example command or experiment snippet to make the body's directives immediately executable.

Suggestions

Trim the Security section's generic AWS guidance (SSE-KMS/TLS/least-privilege explanations) to a one-line pointer to the linked AWS docs, keeping only the skill-specific advice about PII in finding comments and experiment descriptions.

Inline one short end-to-end example — e.g. a single `aws resiliencehubv2` assessment command paired with the FIS experiment that validates its top finding — so the body demonstrates the loop concretely before delegating to the references.

In the monitoring section, the two bullets restating what stop conditions and post-experiment analysis are could be compressed into a single sentence pointing to the Observability skill, saving tokens without losing the lane boundary.

DimensionReasoningScore

Conciseness

The body is a lean router — each section earns its place and nothing explains basics Claude doesn't know, except the Security section, which restates generic AWS guidance (SSE-KMS, `aws:SecureTransport`, least privilege) that could be trimmed to one line plus the links. Efficient with minor instances of over-explanation, matching the 4 anchor.

4 / 5

Actionability

Directives are concrete and ordered ("Run the experiment, confirm the system recovers within its objectives, then mark resolved"; "create a policy, register your service, run an assessment") and the hallucination-prone API details are delegated to a real, well-signaled reference file. It stops short of fully copy-paste-ready guidance in the body itself — no example commands or experiment template inline — so it sits at 4 rather than 5.

4 / 5

Workflow Clarity

The Define → Test → Operate sequence is stated up front and the body enforces an explicit validation checkpoint ("You MUST validate each remediation with an experiment... BEFORE marking the finding resolved") with stop-condition guidance for destructive FIS runs. The full step-by-step SOP lives one reference deep, so checkpoints are present but the detailed sequence with recovery loops is not in the body itself — the 4 anchor fits better than 5.

4 / 5

Progressive Disclosure

A clear overview with three one-level-deep references (lifecycle-workflow.md, best-practices.md, api-reference.md — all present on disk, none nesting further), each introduced with a purpose signal ("For operational patterns and policy design guidance", "READ FIRST before producing any AWS CLI command"), plus a guardrail explaining how to load the files under MCP vs local install. This matches the top anchor.

5 / 5

Total

17

/

20

Passed

Description

92%Weight 40%Scale 1-5

Based on the skill's description, can an agent find and select it at the right time? Clear, specific descriptions lead to better discovery.

A strong description: it states the domain and concrete workflow stages in third person, gives explicit positive and negative trigger conditions in natural user language, and draws sharp boundaries against sibling skills. The only weakness is modest synonym coverage for how users commonly phrase resilience testing.

Suggestions

Add one or two natural synonyms users say for this domain — e.g. "chaos engineering" or "resilience testing / game days" — to the trigger clause so the skill fires on those phrasings too.

DimensionReasoningScore

Specificity

Enumerates concrete actions across the full lifecycle — "from policy creation through failure mode assessment, to FIS experiment validation, to ARC operational controls" — covering all three services comprehensively with no obvious gaps. This matches the comprehensive-coverage anchor rather than the minor-gaps anchor at 4.

5 / 5

Completeness

Explicitly answers what ("Guides the end-to-end AWS resilience lifecycle... Covers the Define → Test → Operate workflow") and when ("Applicable when the user wants a complete resilience strategy... Also applicable for... Not applicable for..."), with concrete trigger phrases and explicit exclusions. Clearly the top anchor.

5 / 5

Trigger Term Quality

Good natural-language triggers ("complete resilience strategy", "planning a resilience program", whether they are "done" after resolving findings) plus service names and CLI namespaces. A few common synonyms users would actually say — "chaos engineering", "game day", "resilience testing" — are absent, so it falls just short of the comprehensive-synonyms anchor at 5.

4 / 5

Distinctiveness Conflict Risk

Claims a clear niche (the integrated three-service lifecycle) and actively disambiguates from sibling skills: "Not applicable for resolving or remediating a specific individual finding (see resilience-hub-failure-mode-assessment), or when a single service is explicitly named". Minimal conflict risk.

5 / 5

Total

19

/

20

Passed

Validation

93%

Checks the skill against the spec for correct structure and formatting. All validation checks must pass before discovery and implementation can be scored.

Validation — 15 / 16 Passed

Validation for skill structure

CriteriaDescriptionResult

frontmatter_unknown_keys

Unknown frontmatter key(s) found; consider removing or moving to metadata

Warning

Total

15

/

16

Passed

Repository
aws/agent-toolkit-for-aws
Reviewed

Table of Contents

Is this your skill?

If you maintain this skill, you can claim it as your own. Once claimed, you can manage eval scenarios, bundle related skills, attach documentation or rules, and ensure cross-agent compatibility.