CtrlK
BlogDocsLog inGet started
Tessl Logo

e2e-tests-studio

REQUIRED when modifying any file in packages/playground-ui or packages/playground. Triggers on: React component creation/modification/refactoring, UI changes, new playground features, bug fixes affecting studio UI. Generates Playwright E2E tests that validate PRODUCT BEHAVIOR, not just UI states.

64

Quality

76%

Does it follow best practices?

Run evals on this skill

Adds up to 20 points to the overall score

View guide

SecuritybySnyk

Passed

No findings from the security scan

Fix and improve this skill with Tessl

tessl review fix ./.claude/skills/e2e-tests-studio/SKILL.md
SKILL.md
Quality
Evals
Security

Quality

Content

71%Weight 40%Scale 1-5

Reviews the quality of instructions and guidance provided to agents. Good implementation is clear, handles edge cases, and produces reliable results.

A highly actionable, well-sequenced skill body with executable code throughout and a clear validation checklist. Its weaknesses are length and duplication (the anti-pattern content and BDD rules appear multiple times) and the absence of any progressive disclosure — the six-pattern library and fixture guidance should live in reference files rather than the main SKILL.md.

Suggestions

Move the six behavior-test patterns and the fixture authoring guide (Steps 4 and 6) into references/ files (e.g., references/test-patterns.md, references/fixtures.md), keeping only one or two exemplar patterns inline in SKILL.md.

De-duplicate the anti-pattern guidance: the 'What NOT to test' list at the top and the 'Anti-Patterns to Avoid' table at the bottom repeat the same content — keep one, and state the BDD nesting rules once instead of across the example, rules list, and template.

Add an explicit feedback loop after Step 7: when a test fails, instruct Claude to read the failure, fix the test or product code, and re-run "pnpm test:e2e" until green before declaring completion.

DimensionReasoningScore

Conciseness

The ~450-line body is concrete and mostly free of filler, but the "What NOT to test" ❌ list duplicates the closing Anti-Patterns table, the BDD nesting rules are stated three times (example, rules list, template), and six full code patterns inflate the token budget. It is more than 'minor instances of over-explanation', so it sits at 3 rather than 4.

3 / 5

Actionability

Everything is executable: copy-paste-ready Playwright specs, exact shell commands ("pnpm build:cli", "cd packages/playground && pnpm test:e2e"), a feature-to-test mapping table, and a pre-completion quality checklist. The six patterns cover the common cases exactly as the top anchor requires.

5 / 5

Workflow Clarity

Steps 1–7 are clearly sequenced with build/start commands and an explicit validation step (Step 7 run + quality checklist). It falls short of 5 because there is no error-recovery feedback loop — nothing tells Claude what to do when a test fails or how to iterate.

4 / 5

Progressive Disclosure

There are no bundle files (references/, scripts/, assets/ absent) and everything is inlined, including ~250 lines of pattern examples and the fixture guide that clearly belong in one-level-deep reference files. Section headers and the quick-reference table provide real structure, so it is above the 'minimal structure' anchor of 2 but below the well-split anchor of 4.

3 / 5

Total

15

/

20

Passed

Description

82%Weight 40%Scale 1-5

Based on the skill's description, can an agent find and select it at the right time? Clear, specific descriptions lead to better discovery.

A strong description: it explicitly states both what the skill does and when it must trigger, with concrete, package-scoped trigger phrases that minimize conflict risk. The only soft spot is that the capability statement is a single action rather than a list of several, and a few natural synonyms (E2E, end-to-end, Playwright) are absent from the trigger terms.

DimensionReasoningScore

Specificity

Names one concrete capability — "Generates Playwright E2E tests that validate PRODUCT BEHAVIOR, not just UI states" — with a specific framework, but does not list several distinct actions. It sits at 'domain plus 1-2 concrete actions', below the 'several specific actions' anchor of 4.

3 / 5

Completeness

Explicitly answers both questions: what ("Generates Playwright E2E tests that validate PRODUCT BEHAVIOR, not just UI states") and when ("REQUIRED when modifying any file in packages/playground-ui or packages/playground. Triggers on: …") with concrete trigger phrases. It clearly matches the top anchor.

5 / 5

Trigger Term Quality

"React component creation/modification/refactoring, UI changes, new playground features, bug fixes affecting studio UI" provides good natural keyword coverage a user would plausibly say. A few natural synonyms (e.g., "E2E", "end-to-end", "Playwright", "playground-ui") are missing, so it falls short of the comprehensive anchor of 5.

4 / 5

Distinctiveness Conflict Risk

The trigger is pinned to specific monorepo packages ("packages/playground-ui or packages/playground") and studio-UI-specific events, giving it a clear niche with minimal overlap risk against generic frontend or testing skills.

5 / 5

Total

17

/

20

Passed

Validation

93%

Checks the skill against the spec for correct structure and formatting. All validation checks must pass before discovery and implementation can be scored.

Validation — 15 / 16 Passed

Validation for skill structure

CriteriaDescriptionResult

frontmatter_unknown_keys

Unknown frontmatter key(s) found; consider removing or moving to metadata

Warning

Total

15

/

16

Passed

Repository
mastra-ai/mastra
Reviewed

Table of Contents

Is this your skill?

If you maintain this skill, you can claim it as your own. Once claimed, you can manage eval scenarios, bundle related skills, attach documentation or rules, and ensure cross-agent compatibility.