CtrlK
BlogDocsLog inGet started
Tessl Logo

e2e-testing

Playwright E2E testing patterns, Page Object Model, configuration, CI/CD integration, artifact management, and flaky test strategies. Use when writing Playwright tests, structuring page objects, or fixing flaky E2E runs in CI.

67

Quality

80%

Does it follow best practices?

Run evals on this skill

Adds up to 20 points to the overall score

View guide

SecuritybySnyk

Passed

No findings from the security scan

Fix and improve this skill with Tessl

tessl review fix ./.agents/skills/e2e-testing/SKILL.md

The canonical home for this skill is e2e-testing in affaan-m/ECC

SKILL.md
Quality
Evals
Security

Quality

Content

72%Weight 40%Scale 1-5

Reviews the quality of instructions and guidance provided to agents. Good implementation is clear, handles edge cases, and produces reliable results.

A highly actionable, lean pattern reference with copy-paste-ready examples throughout. Its weaknesses are the absence of any sequenced workflow with validation checkpoints and the monolithic inline structure that keeps specialized topics in SKILL.md instead of one-level-deep reference files.

Suggestions

Split specialized topics (Wallet/Web3 testing, Financial/Critical flows, CI/CD workflow, Test Report Template) into one-level-deep reference files (e.g. references/ci-cd.md, references/web3.md) linked from short SKILL.md sections.

Add an explicit ordered workflow for flaky tests with validation checkpoints, e.g. 1) reproduce with --repeat-each=10, 2) confirm the cause is fixed by re-running until N consecutive passes, 3) only then remove the quarantine/skip marker.

Fix the mismatch between the shown test layout (which has no pages/ directory) and the page-object import path '../../pages/ItemsPage', and trim boilerplate like the full reporter list to reduce token cost.

DimensionReasoningScore

Conciseness

The body is almost entirely executable code with minimal prose and no explanations of concepts Claude already knows; minor trimmable spots remain, such as the full config boilerplate and the report template. Anchor 5 would require every section to be irreducible, which the boilerplate blocks.

4 / 5

Actionability

Every section contains copy-paste-ready TypeScript, Bash, or YAML with concrete selectors and commands (e.g. 'npx playwright test tests/search.spec.ts --repeat-each=10'), and the bad/good pairs in the flaky section cover the common cases; the only blemish is the '../../pages/ItemsPage' import not matching the shown directory layout.

5 / 5

Workflow Clarity

The content is organized as a pattern reference rather than a sequenced workflow; the flaky-test section implies a sequence (quarantine, then identify via --repeat-each, then fix) but no explicit validation checkpoints or feedback loops are stated. Not 4 because checkpoints are absent rather than minor-gapped; not 2 because sections are individually coherent and the flaky sequence is discernible.

3 / 5

Progressive Disclosure

Sections are well-organized under clear headers, but the ~320-line body is entirely inline with no reference files, and content that would fit one-level-deep references (CI/CD workflow, wallet/web3 testing, financial flows, report template) is inlined in SKILL.md. No bundle files exist to offset this.

3 / 5

Total

15

/

20

Passed

Description

87%Weight 40%Scale 1-5

Based on the skill's description, can an agent find and select it at the right time? Clear, specific descriptions lead to better discovery.

A strong description that clearly states capability areas and includes an explicit, natural trigger clause. Its only weakness is listing topics as nouns rather than concrete actions and omitting a few common synonyms like 'end-to-end'.

DimensionReasoningScore

Specificity

The description names the domain and six specific areas ('Page Object Model, configuration, CI/CD integration, artifact management, and flaky test strategies'), going well beyond the 1-2 actions of anchor 3, but they are topic nouns rather than the concrete verb-phrased actions of the anchor-5 example.

4 / 5

Completeness

Both questions are explicitly answered: a clear 'what' (the six capability areas) and a concrete 'when' via an explicit 'Use when...' clause with three distinct triggers, matching the anchor-5 example.

5 / 5

Trigger Term Quality

'Use when writing Playwright tests, structuring page objects, or fixing flaky E2E runs in CI' contains natural phrases users would actually say, but common synonyms like 'end-to-end testing' or 'browser tests' are missing.

4 / 5

Distinctiveness Conflict Risk

'Playwright' is named throughout, giving the description a clear niche with distinct triggers and minimal risk of firing for unrelated testing skills.

5 / 5

Total

18

/

20

Passed

Validation

100%

Checks the skill against the spec for correct structure and formatting. All validation checks must pass before discovery and implementation can be scored.

Validation — 16 / 16 Passed

Validation for skill structure

No warnings or errors.

Repository
affaan-m/ECC
Reviewed

Table of Contents

Is this your skill?

If you maintain this skill, you can claim it as your own. Once claimed, you can manage eval scenarios, bundle related skills, attach documentation or rules, and ensure cross-agent compatibility.