CtrlK
BlogDocsLog inGet started
Tessl Logo

e2e-testing

Playwright E2E testing patterns, Page Object Model, configuration, CI/CD integration, artifact management, and flaky test strategies.

56

Quality

65%

Does it follow best practices?

Run evals on this skill

Adds up to 20 points to the overall score

View guide

SecuritybySnyk

Medium

Suggest reviewing before use

Fix and improve this skill with Tessl

tessl review fix ./.kiro/skills/e2e-testing/SKILL.md

The canonical home for this skill is e2e-testing in affaan-m/ECC

SKILL.md
Quality
Evals
Security

Quality

Content

65%Weight 40%Scale 1-5

Reviews the quality of instructions and guidance provided to agents. Good implementation is clear, handles edge cases, and produces reliable results.

A lean, highly actionable pattern reference built almost entirely from executable Playwright code, with strong bad/good fix examples for flakiness. Its weaknesses are structural: it is a single monolithic file with no reference files for niche topics, and its flaky-test diagnostic flow lacks an explicit re-run validation checkpoint.

Suggestions

Split niche sections (Wallet/Web3 Testing, Financial/Critical Flow Testing, Test Report Template) into one-level-deep reference files under references/ and link them from SKILL.md.

Close the flaky-test workflow loop with an explicit validation step, e.g. 'Re-run with --repeat-each=10 to confirm the fix; if still flaky, quarantine with test.fixme and file an issue.'

Fix the minor correctness gaps: add the pages/ directory to the file-organization tree so the POM import path resolves, and replace the invalid 'videosPath' option with Playwright's real video output configuration.

DimensionReasoningScore

Conciseness

The body is almost entirely executable code with virtually no concept-explanation padding — it assumes Claude already knows how Playwright works. It falls just short of anchor 5 because a few sections do not fully earn their tokens: the Test Report Template and the wallet/web3 and financial-testing sections are niche, and the config block repeats 'networkidle' waits throughout. Clearly above anchor 3's 'some unnecessary explanation'.

4 / 5

Actionability

Nearly all guidance is complete, copy-paste-ready TypeScript/YAML/bash covering the common cases (POM class, spec structure, defineConfig, CI workflow, flaky-test commands). Minor gaps keep it below anchor 5: the test imports '../../pages/ItemsPage' but the shown directory tree has no pages/ directory, and the video snippet uses a non-existent top-level 'videosPath' option rather than Playwright's actual output-dir configuration.

4 / 5

Workflow Clarity

This is a patterns catalog rather than a sequenced process; the closest thing to a workflow is the flaky-test section (quarantine → reproduce with --repeat-each → match a common cause → apply fix), which is a rough sequence. It matches anchor 3 ('sequence present but checkpoints missing or implicit'): there is no explicit validation step telling the reader to re-run the repeated test to confirm the fix, and the other sections have no step ordering at all.

3 / 5

Progressive Disclosure

There are no bundle files at all (no references/, scripts/, or assets/), so all ~320 lines live inline in SKILL.md. Section headers are clear, but content that belongs in one-level-deep reference files — the report template, wallet/web3 testing, and financial-flow testing sections — is inlined. This matches anchor 3 ('some structure... content that should be separate is inline'), above anchor 2 because organization is genuinely good, below anchor 4 because nothing is split out.

3 / 5

Total

14

/

20

Passed

Description

65%Weight 40%Scale 1-5

Based on the skill's description, can an agent find and select it at the right time? Clear, specific descriptions lead to better discovery.

A domain-specific, reasonably distinct description that names concrete topic areas but omits any explicit trigger guidance ('Use when...') and uses noun-phrase listing instead of action verbs. The missing 'when' clause is the primary gap, capping completeness and leaving users without natural activation cues.

Suggestions

Add an explicit trigger clause, e.g. 'Use when writing or debugging Playwright end-to-end tests, setting up test CI pipelines, or fixing flaky E2E tests.'

Convert noun topics into concrete actions (e.g. 'Build Page Object Models, configure Playwright projects, quarantine and diagnose flaky tests').

Include common user phrasings and synonyms such as 'end-to-end tests', 'integration tests', and 'test suite' to improve trigger-term coverage.

DimensionReasoningScore

Specificity

The description names a clear domain ('Playwright E2E testing') and several concrete topic areas ('Page Object Model', 'configuration, CI/CD integration, artifact management, and flaky test strategies'), but everything is phrased as nouns rather than actions — it describes subject matter, not what the skill does. This sits between anchor 2 (domain named, actions minimal) and anchor 4 (several specific actions listed); the topics are specific but the verb-less framing keeps it below 4.

3 / 5

Completeness

The 'what' is clear (a list of Playwright testing pattern areas), but there is no 'when' — no 'Use when...' clause or equivalent explicit trigger guidance anywhere in the description. Per the rubric guideline, a missing 'Use when' clause caps completeness at 3, which matches anchor 3 exactly ('clear what, when missing or only weakly implied').

3 / 5

Trigger Term Quality

Natural keywords users would actually say are present: 'Playwright', 'E2E testing', 'Page Object Model', 'CI/CD', 'flaky test'. A few common variations are missing — 'end-to-end' spelled out, 'integration testing', 'test suite' — so it does not reach the comprehensive synonym coverage of anchor 5 but clearly exceeds the partial coverage of anchor 3.

4 / 5

Distinctiveness Conflict Risk

'Playwright E2E testing patterns, Page Object Model' carves out a clear, distinct niche with specific tooling terms unlikely to collide with other skills (no generic testing or QA skill would match these triggers better). This matches anchor 5 ('clear niche with distinct triggers; minimal conflict risk'); it is clearly above anchor 4's 'minor overlap risk' framing since the named tool and patterns are unambiguous.

5 / 5

Total

15

/

20

Passed

Validation

93%

Checks the skill against the spec for correct structure and formatting. All validation checks must pass before discovery and implementation can be scored.

Validation — 15 / 16 Passed

Validation for skill structure

CriteriaDescriptionResult

metadata_version

'metadata.version' is missing

Warning

Total

15

/

16

Passed

Repository
affaan-m/ECC
Reviewed

Table of Contents

Is this your skill?

If you maintain this skill, you can claim it as your own. Once claimed, you can manage eval scenarios, bundle related skills, attach documentation or rules, and ensure cross-agent compatibility.