CtrlK
BlogDocsLog inGet started
Tessl Logo

resilience-hub-failure-mode-assessment

Runs and interprets AWS Resilience Hub v2 failure mode assessments. Covers starting assessments, understanding findings (severity, categories, recommendations), triaging by achievability, working with AI-generated service functions, and resolving findings. Applies when the user wants to run an assessment, review findings, or understand failure modes, or has a specific finding and asks how to resolve, remediate, or fix it. Does not apply to initial setup (use resilience-hub-getting-started) or FIS experiments.

72

Quality

88%

Does it follow best practices?

Run evals on this skill

Adds up to 20 points to the overall score

View guide

SecuritybySnyk

Low

Low-risk findings worth noting

SKILL.md
Quality
Evals
Security

Quality

Content

85%Weight 40%Scale 1-5

Reviews the quality of instructions and guidance provided to agents. Good implementation is clear, handles edge cases, and produces reliable results.

A well-structured overview body that appropriately delegates the detailed SOP to a single, clearly signaled reference containing executable commands, validation checkpoints, and troubleshooting. The body itself stays lean with concrete commands and decision rules; the only room for improvement is minor tightening of the MCP/local-install guardrail and security sections and showing full command flags in the body.

Suggestions

Merge the top-level MCP note into the guardrail section (or shorten it to a pointer) — both cover the same MCP-optional point, and the guardrail's retrieve_skill instructions could state the two load modes once instead of repeating the file-fetch rule.

In 'AI-generated service functions are wrong', show the update-service-function call with its key flags (--service-arn, --service-function-id, --name, --criticality) so the body's most likely ad-hoc command is copy-paste ready without opening the reference.

DimensionReasoningScore

Conciseness

The body is lean and assumes competence — no explanation of concepts Claude already knows, and troubleshooting entries are direct (e.g., 'Verify the invoker role (and any cross-account roles) can describe resources in all configured regions'). It is not a 5 because the MCP-vs-local guardrail section and the top-level MCP note have some redundancy and the security section could be trimmed slightly. It is clearly above 3, which would require unnecessary explanation or padding.

4 / 5

Actionability

Concrete commands and API operations appear throughout — 'aws resiliencehubv2 update-service-function', 'create-service-function-resources', achievability from 'get-service / list-failure-mode-assessments' — and the full SOP with copy-paste commands and examples lives in the reference. It is not a 5 because some body-level commands are shown without their key flags (e.g., --service-arn/--service-function-id for update-service-function), leaving minor gaps if read in isolation. It is not a 3 because nothing is pseudocode or abstract.

4 / 5

Workflow Clarity

The body directs 'follow the procedure exactly' to a well-signaled reference containing a 9-step sequence with explicit validation checkpoints — dependency verification before anything runs, explicit cost confirmation before the billable assessment, polling 'until status is SUCCESS or FAILED', and error-code handling on report generation. It is not a 4 because the delegated procedure includes complete validation and failure-recovery loops rather than just most checkpoints.

5 / 5

Progressive Disclosure

The body is a concise overview with one clearly signaled, one-level-deep reference (references/assessment-workflow.md, which exists in the bundle) linked contextually from the run/interpret section and both troubleshooting entries. It is not a 4 because there is no content that should have been split out left inline, no buried references, and no nesting beyond one level.

5 / 5

Total

18

/

20

Passed

Description

92%Weight 40%Scale 1-5

Based on the skill's description, can an agent find and select it at the right time? Clear, specific descriptions lead to better discovery.

A strong description: third-person voice, concrete capability list, explicit 'Applies when' triggers with synonyms, and explicit non-goals that route to sibling skills. The only minor gap is that a few natural user phrasings (e.g., resiliency review, reliability analysis) are not covered.

DimensionReasoningScore

Specificity

The description enumerates concrete, distinct actions — 'starting assessments, understanding findings (severity, categories, recommendations), triaging by achievability, working with AI-generated service functions, and resolving findings' — comprehensively covering the skill's capabilities, matching the score-5 anchor of multiple specific actions with comprehensive coverage. It is not a 4 because the action list leaves no notable coverage gaps within the stated scope.

5 / 5

Completeness

Both 'what' (running and interpreting failure mode assessments, triaging, resolving findings) and 'when' ('Applies when the user wants to run an assessment... or has a specific finding and asks how to resolve, remediate, or fix it') are explicit with concrete trigger phrases, plus explicit exclusions. It is not a 4 because the 'when' clause is already fully explicit with multiple concrete trigger scenarios.

5 / 5

Trigger Term Quality

Trigger phrases like 'run an assessment, review findings, or understand failure modes, or has a specific finding and asks how to resolve, remediate, or fix it' are natural user language with good synonym coverage (resolve/remediate/fix). It falls short of 5 because it omits some natural variations users might say (e.g., 'resiliency review', 'weaknesses', 'chaos/reliability analysis') while 'AI-generated service functions' is more jargon than a user trigger.

4 / 5

Distinctiveness Conflict Risk

A clear niche (Resilience Hub v2 failure mode assessments) is reinforced by explicit boundary guidance — 'Does not apply to initial setup (use resilience-hub-getting-started) or FIS experiments' — routing adjacent use cases away. It is not a 4 because the explicit cross-references to sibling skills leave minimal overlap risk.

5 / 5

Total

19

/

20

Passed

Validation

93%

Checks the skill against the spec for correct structure and formatting. All validation checks must pass before discovery and implementation can be scored.

Validation — 15 / 16 Passed

Validation for skill structure

CriteriaDescriptionResult

frontmatter_unknown_keys

Unknown frontmatter key(s) found; consider removing or moving to metadata

Warning

Total

15

/

16

Passed

Repository
aws/agent-toolkit-for-aws
Reviewed

Table of Contents

Is this your skill?

If you maintain this skill, you can claim it as your own. Once claimed, you can manage eval scenarios, bundle related skills, attach documentation or rules, and ensure cross-agent compatibility.