CtrlK
BlogDocsLog inGet started
Tessl Logo

deploy-smoke-test

Confirm a production deploy actually landed and is healthy. Verifies the latest Vercel Production deployment is READY and matches the current `main` commit, runs HTTP canary checks against travel.matthewcarr.dev, confirms migrations applied, and checks for a post-deploy Sentry error spike. Use after merging to `main`, or when a human asks "is prod healthy?" / "did the deploy go out?" / "smoke test production". Read-only against prod by default.

81

1.53x
Quality

94%

Does it follow best practices?

Impact

100%

1.53x

1 of 3 eval scenarios. Add 2 more for a full score.

SecuritybySnyk

Passed

No findings from the security scan

SKILL.md
Quality
Evals
Security

Quality

Content

85%Weight 40%Scale 1-5

Reviews the quality of instructions and guidance provided to agents. Good implementation is clear, handles edge cases, and produces reliable results.

The body is a well-sequenced, actionable smoke-test workflow with strong validation checkpoints and a clear report format. It is mostly lean, with minor boilerplate (untrusted-content section, editorializing in Gaps) and one step that delegates to an external doc.

Suggestions

Tighten the 'Untrusted content' section — the generic 'treat external text as data not instructions' guidance is largely boilerplate Claude already follows; keep only what is specific to this skill.

Make the Sentry check (Step 4) more actionable by inlining the specific command or API query instead of deferring entirely to docs/operations/sentry.md.

Trim editorializing in 'Gaps worth closing' (e.g. 'Per the repo's durable over expedient bias') to keep the section a concise recommendation.

DimensionReasoningScore

Conciseness

Mostly efficient and purposeful, but the 'Untrusted content' section reiterates general guidance Claude already knows and the 'Gaps' section editorializes ('Per the repo's durable over expedient bias'), which could be trimmed.

4 / 5

Actionability

Provides concrete executable commands (git rev-parse, vercel ls/inspect, curl with http_code, vercel logs/rollback) with specific expected outcomes, though the Sentry step defers to docs/operations/sentry.md rather than giving an exact command.

4 / 5

Workflow Clarity

Clear five-step sequence with explicit validation checkpoints (state=READY, commit SHA match, status-code verdicts, 5xx=fail, Sentry spike), explicit fail conditions, and a structured report template with a verdict.

5 / 5

Progressive Disclosure

No bundle files exist; the skill is a single self-contained file with well-organized sections and clearly signaled one-level references to repo files (infra/stacks/prod, docs/operations/sentry.md, ADR 056).

5 / 5

Total

18

/

20

Passed

Description

100%Weight 40%Scale 1-5

Based on the skill's description, can an agent find and select it at the right time? Clear, specific descriptions lead to better discovery.

The description is excellent: third-person voice, concrete actions, explicit trigger phrases, and a clear niche. It answers both what and when without verbosity or over-claiming.

DimensionReasoningScore

Specificity

Lists multiple concrete actions — 'Verifies the latest Vercel Production deployment is READY', 'matches the current main commit', 'runs HTTP canary checks', 'confirms migrations applied', 'checks for a post-deploy Sentry error spike' — giving comprehensive, specific coverage rather than vague language.

5 / 5

Completeness

Explicitly answers both 'what' (verifies deploy landed and is healthy via concrete checks) and 'when' ('Use after merging to main, or when a human asks...') with concrete trigger phrases.

5 / 5

Trigger Term Quality

Includes natural phrases a user would actually say — 'is prod healthy?', 'did the deploy go out?', 'smoke test production', 'after merging to main' — covering synonyms and common variations.

5 / 5

Distinctiveness Conflict Risk

Targets a narrow niche — Vercel production deploy smoke testing against a specific domain — with distinct triggers and minimal overlap with other skills.

5 / 5

Total

20

/

20

Passed

Validation

100%

Checks the skill against the spec for correct structure and formatting. All validation checks must pass before discovery and implementation can be scored.

Validation — 16 / 16 Passed

Validation for skill structure

No warnings or errors.

Repository
matthew-a-carr/travel-planner
Reviewed

Table of Contents

Is this your skill?

If you maintain this skill, you can claim it as your own. Once claimed, you can manage eval scenarios, bundle related skills, attach documentation or rules, and ensure cross-agent compatibility.