CtrlK
BlogDocsLog inGet started
Tessl Logo

troubleshooting-application-failures

Troubleshoots failing applications by discovering and analyzing CloudWatch log groups to identify error patterns, root causes, and actionable solutions. Use when an application is experiencing failures and log-based diagnosis is needed.

68

Quality

83%

Does it follow best practices?

Run evals on this skill

Adds up to 20 points to the overall score

View guide

SecuritybySnyk

Low

Low-risk findings worth noting

SKILL.md
Quality
Evals
Security

Quality

Content

92%Weight 40%Scale 1-5

Reviews the quality of instructions and guidance provided to agents. Good implementation is clear, handles edge cases, and produces reliable results.

The body exemplifies progressive disclosure: a lean, non-padded overview that delegates the multi-step procedure to a single real reference file, which is well-sequenced with explicit validation gates and error-recovery loops. The only gap keeping actionability from the top anchor is that the body itself carries no commands or examples, relying entirely on the referenced procedure for executable detail.

DimensionReasoningScore

Conciseness

The ~35-line body is lean and assumes Claude's competence: no explanation of what CloudWatch is or how logs work, no padded sections, and every part (overview, procedure pointer, three short failure-mode fixes) earns its place, matching the "lean and efficient" anchor. It is above anchor 4 because there is no over-explanation to trim — the brief Overview is orientation, not padding.

5 / 5

Actionability

The body gives mostly concrete guidance: exact log group name patterns ("/aws/lambda/function-name"), exact required IAM permissions ("logs:DescribeLogGroups", "logs:StartQuery", "logs:GetQueryResults"), and a specific fix for timeouts ("Reduce the time window or limit results"). However, the executable procedure itself is deferred to the reference file and the body contains no commands or examples of its own, so it stops short of the fully copy-paste-ready anchor 5; it is clearly above anchor 3, which requires missing key details or pseudocode.

4 / 5

Workflow Clarity

The body clearly signals "follow the procedure exactly" via a real, one-level-deep reference whose verified content is a 9-step sequence with explicit validation gates ("Do not proceed until you have received and confirmed all required parameters", "Verify Dependencies", "Validate Log Groups and Check Availability", "only proceed with log analysis if log streams were found") and a wait-for-query-results step. The body's own Troubleshooting section adds error-recovery paths (no log groups → ask user; access denied → check permissions; timeout → narrow the query), matching the anchor-5 pattern of clear sequence with explicit validation and feedback loops; the operations are read-only, so the destructive/batch cap does not apply.

5 / 5

Progressive Disclosure

The body is a clear overview that appropriately splits content: the SOP lives in references/application-failure-troubleshooting.md (verified to exist, 350 lines, no nested references — one level deep), signaled with a clean markdown link under a task-oriented heading, while short failure-mode guidance stays inline. This matches the anchor-5 example of a concise overview with well-signaled one-level-deep references and easy navigation.

5 / 5

Total

19

/

20

Passed

Description

75%Weight 40%Scale 1-5

Based on the skill's description, can an agent find and select it at the right time? Clear, specific descriptions lead to better discovery.

A solid description in third-person voice with concrete domain actions and an explicit "Use when" clause covering both what and when. It scores consistently at anchor 4 across dimensions: specific and distinct, but with slightly generic outcome phrasing, a narrow when-clause, and a few missing natural trigger synonyms that keep it from anchor 5.

Suggestions

Replace the generic "root causes, and actionable solutions" tail with concrete capabilities, e.g., "...search for error patterns and stack traces, determine root causes, and generate prioritized remediation recommendations".

Broaden the "Use when" clause with natural trigger phrases users would actually say, e.g., "Use when an application is failing, throwing errors or 5xx responses, crashing, or when log-based diagnosis of CloudWatch logs is needed".

Add domain keywords such as "stack traces", "exceptions", or "production issues" to improve trigger-term coverage and further distinguish the skill from general debugging skills.

DimensionReasoningScore

Specificity

"Troubleshoots failing applications by discovering and analyzing CloudWatch log groups to identify error patterns, root causes, and actionable solutions" lists several concrete actions (discover, analyze, identify patterns/causes), but "root causes and actionable solutions" is outcome-descriptive and slightly generic rather than a fully concrete capability, so it falls just below the comprehensive anchor.

4 / 5

Completeness

It clearly answers "what" (discovers/analyzes CloudWatch log groups, identifies error patterns and root causes) and has an explicit "Use when an application is experiencing failures and log-based diagnosis is needed" clause, so both are present; the when-clause is a single formal condition rather than a set of concrete trigger phrases, which is the anchor-4 pattern ("when could be more explicit or specific"). It is clearly above anchor 3, where "when" is missing or only weakly implied.

4 / 5

Trigger Term Quality

Natural terms like "troubleshoot", "failing applications", "failures", "CloudWatch", and "log-based diagnosis" give good coverage an AWS user would plausibly say, but common variations such as "errors", "crashes", "exceptions", "stack traces", or "production issues" are missing, matching the "good keyword coverage; a few natural terms missing" anchor rather than 5.

4 / 5

Distinctiveness Conflict Risk

The CloudWatch log-group niche is mostly distinct with minimal conflict risk against unrelated skills, but the broad opening "Troubleshoots failing applications" could overlap a general debugging or log-analysis skill, placing it at "mostly distinct; minor overlap risk with closely related skills" rather than the clear-niche anchor 5.

4 / 5

Total

16

/

20

Passed

Validation

93%

Checks the skill against the spec for correct structure and formatting. All validation checks must pass before discovery and implementation can be scored.

Validation — 15 / 16 Passed

Validation for skill structure

CriteriaDescriptionResult

frontmatter_unknown_keys

Unknown frontmatter key(s) found; consider removing or moving to metadata

Warning

Total

15

/

16

Passed

Repository
aws/agent-toolkit-for-aws
Reviewed

Table of Contents

Is this your skill?

If you maintain this skill, you can claim it as your own. Once claimed, you can manage eval scenarios, bundle related skills, attach documentation or rules, and ensure cross-agent compatibility.