CtrlK
BlogDocsLog inGet started
Tessl Logo

running-release-tests

Run automated release testing (UI or API) via the AWS DevOps Agent using a pre-configured test profile. Use when the user wants to validate multi-step workflows, verify features, check for regressions, or test API endpoints. Trigger words include run tests, UAT, test my app, test profile, UI test, API test, automated testing, regression test, QA, end-to-end test, run the QA agent.

72

Quality

88%

Does it follow best practices?

Run evals on this skill

Adds up to 20 points to the overall score

View guide

SecuritybySnyk

Passed

No findings from the security scan

SKILL.md
Quality
Evals
Security

Quality

Content

88%Weight 40%Scale 1-5

Reviews the quality of instructions and guidance provided to agents. Good implementation is clear, handles edge cases, and produces reliable results.

Highly actionable operational content: complete tool-call and CLI syntax, explicit sequencing with user-interaction checkpoints, polling loops, and a thorough error-handling section. The main improvements are structural — deduplicating the agent-space and results-presentation steps shared by the primary and fallback paths, and optionally splitting the fallback path into its own reference file.

Suggestions

Factor the duplicated "Select Agent Space" and "Present Results" steps out of the two paths (e.g. a shared "Common steps" note) to remove roughly 30 lines of repetition.

Consider moving the aws-mcp fallback path to a references/fallback-cli.md file and signaling it from the main workflow, keeping SKILL.md focused on the primary tool-based path.

Add one line on what to do if report retrieval (get_release_ui_testing_report / get_release_api_testing_report) fails or returns empty, closing the only validation gap in the workflow.

DimensionReasoningScore

Conciseness

The body is command-first with no explanation of concepts Claude already knows, but the "Select Agent Space" step and the "Present Results" step are duplicated nearly verbatim between the primary and fallback paths and could be factored out, fitting anchor 4 (minor trimming possible) rather than 5.

4 / 5

Actionability

Fully executable throughout: concrete tool calls with parameters (create_release_testing_job(test_profile_id=..., webhook_event_message=...)) including expected responses, complete CLI commands with all flags, exact poll intervals (30s/20s), and a concrete report filename pattern — matching anchor 5.

5 / 5

Workflow Clarity

Clear numbered sequence with explicit validation checkpoints ("Do NOT proceed until the user has selected one", "You MUST wait for the user to respond") and feedback loops for error recovery (throttle retry up to 3 times, cancel after 5 minutes without IN_PROGRESS, credential-refresh guidance), matching anchor 5; the destructive/batch cap does not apply since this only reads test results.

5 / 5

Progressive Disclosure

No bundle files exist and the ~200-line body is well-sectioned with clear headers and navigation, but it is a monolithic single file where the ~90-line aws-mcp fallback path could arguably live in a separate reference — good structure with minor organization gaps, anchor 4 rather than 5 (the under-50-line exception does not apply).

4 / 5

Total

18

/

20

Passed

Description

87%Weight 40%Scale 1-5

Based on the skill's description, can an agent find and select it at the right time? Clear, specific descriptions lead to better discovery.

A strong description that clearly states what the skill does and when to use it, with an excellent spread of natural trigger terms in third-person voice. Its main weakness is trigger generality: terms like "run tests" and "QA" risk firing on ordinary testing requests that have nothing to do with the AWS DevOps Agent.

Suggestions

Qualify the generic trigger terms with the AWS context, e.g. "run tests (via AWS DevOps Agent)" or replace bare "run tests" with "run release tests", so it does not compete with ordinary unit-testing requests.

Mention the downstream capabilities (monitor live test progress, generate a markdown test report) in the what-clause to lift specificity toward comprehensive coverage.

DimensionReasoningScore

Specificity

"Run automated release testing (UI or API) via the AWS DevOps Agent using a pre-configured test profile" plus "validate multi-step workflows, verify features, check for regressions, or test API endpoints" lists several concrete actions, matching anchor 4 rather than 3; it falls short of 5 because coverage has gaps (no mention of report generation or job monitoring).

4 / 5

Completeness

Explicitly answers both what ("Run automated release testing (UI or API) via the AWS DevOps Agent using a pre-configured test profile") and when ("Use when the user wants to validate multi-step workflows, verify features, check for regressions, or test API endpoints") with concrete trigger phrases, matching anchor 5.

5 / 5

Trigger Term Quality

"run tests, UAT, test my app, test profile, UI test, API test, automated testing, regression test, QA, end-to-end test, run the QA agent" is comprehensive natural-term coverage including synonyms and casual phrasings, matching anchor 5 exactly.

5 / 5

Distinctiveness Conflict Risk

The AWS DevOps Agent and test-profile scoping give it a niche, but very generic triggers like "run tests", "QA", and "automated testing" could overlap with general unit-testing or QA skills, fitting anchor 3 (could still overlap with similar skills) rather than 4 (only minor overlap).

3 / 5

Total

17

/

20

Passed

Validation

100%

Checks the skill against the spec for correct structure and formatting. All validation checks must pass before discovery and implementation can be scored.

Validation — 16 / 16 Passed

Validation for skill structure

No warnings or errors.

Repository
aws/agent-toolkit-for-aws
Reviewed

Table of Contents

Is this your skill?

If you maintain this skill, you can claim it as your own. Once claimed, you can manage eval scenarios, bundle related skills, attach documentation or rules, and ensure cross-agent compatibility.