CtrlK
BlogDocsLog inGet started
Tessl Logo

mastra-smoke-test

Smoke test Mastra projects locally or deploy to staging/production. Tests Studio UI, agents, tools, workflows, traces, memory, and more. Supports both local development and cloud deployments.

60

Quality

71%

Does it follow best practices?

Run evals on this skill

Adds up to 20 points to the overall score

View guide

SecuritybySnyk

Low

Low-risk findings worth noting

Fix and improve this skill with Tessl

tessl review fix ./.claude/skills/mastra-smoke-test/SKILL.md
SKILL.md
Quality
Evals
Security

Quality

Content

73%Weight 40%Scale 1-5

Reviews the quality of instructions and guidance provided to agents. Good implementation is clear, handles edge cases, and produces reliable results.

A strong, executable skill body with excellent workflow gating, validation loops, and a verified, well-organized one-level reference bundle. Its weaknesses are duplicated test/parameter tables, an oversized inline browser-smoke section that belongs in a reference, and an undefined `smoke test` entry command.

Suggestions

Merge the "Test Options (--test)" table into the "Mandatory Test Checklist" table (they map the same options to the same reference files) to remove the duplication.

Move the "Local Studio Browser Smoke" routes/evidence table and report template into a reference file (e.g., references/browser-smoke.md), keeping only the when-to-run rules and the dev-server health-check inline.

Define the `smoke test` entry command — state what tool or script provides it (or replace it with the actual commands to run) so the Usage examples are fully executable.

DimensionReasoningScore

Conciseness

Mostly efficient — tables, concrete commands, and templates rather than prose — but there is real duplication: the "Test Options (--test)" table repeats the "Mandatory Test Checklist" table's option/reference mapping, and "Multi-Environment Support" restates what the Usage examples already show. Not 2 because there is no explanatory padding of concepts Claude already knows; not 4 because the duplicated tables and the ~70-line inline browser section could be meaningfully tightened.

3 / 5

Actionability

Highly executable throughout: copy-paste `gh pr view`/`gh pr list` queries with expected output shape, a curl dev-server restart-and-poll loop, route tables with pass criteria, and report templates. Not 5 because the `smoke test` entry command itself is never defined or sourced — where it comes from (a script, CLI, or harness) is unstated, leaving a gap in the most important command.

4 / 5

Workflow Clarity

Clear sequenced workflow with explicit validation checkpoints and feedback loops: verify the dev server is alive before opening the browser, restart-and-poll for readiness, a fallback when networkidle times out, gating rules ("Do not create the alpha smoke-test project until the automatic alpha publish workflow has completed"), a partial-publish recovery branch, and a mandatory checklist with per-test read-execute-mark progression and explicit blocker/reporting rules. Matches the top anchor including checklist and error-recovery loops.

5 / 5

Progressive Disclosure

Verified against the actual bundle: all referenced paths exist (all 12 references/tests/*.md, every references/*.md named in the body and References table, and the 3 scripts/*.sh), references are one level deep, and the body instructs staying in SKILL.md until the workflow branches. Not 5 because the "Local Studio Browser Smoke" section (routes table, expectations, report template) is detailed inline content that belongs in a reference file, adding ~70 lines to the overview.

4 / 5

Total

16

/

20

Passed

Description

70%Weight 40%Scale 1-5

Based on the skill's description, can an agent find and select it at the right time? Clear, specific descriptions lead to better discovery.

The description is specific and distinctive with good natural trigger terms, but it completely lacks a "when to use" clause, capping completeness at 3. The final sentence is redundant with the first and could be replaced by trigger guidance.

Suggestions

Add an explicit trigger clause, e.g. "Use when the user asks to smoke test a Mastra project, verify a Mastra release (alpha or stable), or test Mastra Studio/agent/workflow functionality before or after deploying to staging or production."

Replace the redundant final sentence ("Supports both local development and cloud deployments") with additional trigger variations such as "verify a Mastra release" or "check Mastra agents/tools/workflows after a deploy".

Clarify the "and more" hedge by naming the remaining tested areas (scorers, MCP, error handling, experiments) or dropping the phrase.

DimensionReasoningScore

Specificity

"Smoke test Mastra projects locally or deploy to staging/production" and "Tests Studio UI, agents, tools, workflows, traces, memory" list several concrete, specific capabilities. Falls short of 5 because "and more" hedges coverage and "Supports both local development and cloud deployments" restates the first sentence rather than adding a capability; it is clearly above 3 (which expects only 1-2 concrete actions).

4 / 5

Completeness

The "what" is clear (smoke tests Mastra projects across the listed feature areas), but there is no "Use when..." clause or any equivalent explicit trigger guidance — "Supports both local development and cloud deployments" describes capability, not when to invoke the skill. Per the rubric guideline, a missing 'Use when' clause caps completeness at 3; it is not 2 because the "what" is concrete and specific.

3 / 5

Trigger Term Quality

Includes natural terms users would say: "smoke test", "Mastra", "deploy", "staging", "production", "Studio UI", "agents", "tools", "workflows", "traces", "memory". Not 5 because there are no synonyms or common variations (e.g., "verify", "check", "regression test") beyond the single "smoke test" phrasing; clearly above 3's "some relevant keywords but missing variations".

4 / 5

Distinctiveness Conflict Risk

Clear niche: "Mastra" plus "Studio UI", "agents", "workflows", "traces" are framework-specific triggers with minimal overlap risk against generic testing or deployment skills. Nothing in the description would plausibly route a non-Mastra task to it.

5 / 5

Total

16

/

20

Passed

Validation

93%

Checks the skill against the spec for correct structure and formatting. All validation checks must pass before discovery and implementation can be scored.

Validation — 15 / 16 Passed

Validation for skill structure

CriteriaDescriptionResult

referenced_paths_exist

Referenced path issues: 14 deeper-than-1-level

Warning

Total

15

/

16

Passed

Repository
mastra-ai/mastra
Reviewed

Table of Contents

Is this your skill?

If you maintain this skill, you can claim it as your own. Once claimed, you can manage eval scenarios, bundle related skills, attach documentation or rules, and ensure cross-agent compatibility.