CtrlK
BlogDocsLog inGet started
Tessl Logo

e2e-testing

End-to-end testing workflow with Playwright for browser automation, visual regression, cross-browser testing, and CI/CD integration.

57

Quality

66%

Does it follow best practices?

Run evals on this skill

Adds up to 20 points to the overall score

View guide

SecuritybySnyk

Passed

No findings from the security scan

Fix and improve this skill with Tessl

tessl review fix ./plugins/AI-Agents-Safe-Coding-Skills/skills/e2e-testing/SKILL.md

The canonical home for this skill is e2e-testing in administrakt0r/AI-Agents-Safe-Coding-Skills

SKILL.md
Quality
Evals
Security

Quality

Content

67%Weight 40%Scale 1-5

Reviews the quality of instructions and guidance provided to agents. Good implementation is clear, handles edge cases, and produces reliable results.

The content is a well-structured, lean workflow orchestrator with clear phase sequencing and a final quality-gates checklist, but it offers only high-level action lists and copy-paste sub-skill invocations rather than concrete executable testing code. Adding per-phase validation and example code or commands would strengthen the weakest dimensions.

Suggestions

Add concrete, executable guidance to each phase (e.g., a sample Playwright test snippet or `npx playwright install` commands) instead of only high-level action verbs and sub-skill invocations.

Insert validation checkpoints between phases (e.g., "Run `npx playwright test` and confirm zero failures before proceeding to cross-browser setup") to add feedback loops.

Move per-phase detail into one-level-deep reference files (e.g., references/visual-regression.md) with clear links from the overview to improve progressive disclosure for this >50-line skill.

DimensionReasoningScore

Conciseness

The body is lean — short numbered action lists, brief copy-paste prompts, and no explanation of concepts Claude already knows — matching the efficient-with-minor-trim anchor. It is not a 5 because the per-phase "Skills to Invoke / Actions / Copy-Paste Prompts" scaffold repeats verbatim across seven phases, adding some structural redundancy.

4 / 5

Actionability

The copy-paste prompts ("Use @playwright-skill to set up Playwright testing") are executable invocations, but the action lists are high-level directives ("Write test scripts", "Add assertions") with no concrete code or commands for the actual testing tasks, matching the some-concrete-but-incomplete anchor. It is not a 4 because the skill relies entirely on named sub-skills rather than providing executable testing code or commands.

3 / 5

Workflow Clarity

Seven phases are explicitly sequenced and each lists ordered actions, with a final Quality Gates checklist serving as a verification checkpoint, matching the clear-sequence-with-most-checkpoints anchor. It is not a 5 because there are no per-phase validation gates or error-recovery feedback loops between phases, only an end-of-workflow checklist.

4 / 5

Progressive Disclosure

The body is well-organized into clearly labeled phases, a Quality Gates section, and Related Workflow Bundles, giving good structure with minor organization gaps. It is not a 5 because all content is inlined in SKILL.md with no one-level-deep file references (no references/ or assets/ bundle), and the per-phase detail could be split out for a >50-line skill.

4 / 5

Total

15

/

20

Passed

Description

66%Weight 40%Scale 1-5

Based on the skill's description, can an agent find and select it at the right time? Clear, specific descriptions lead to better discovery.

The description is specific and well-keyworded with a clear statement of capabilities tied to Playwright, but it omits any explicit "Use when..." trigger guidance, which caps its completeness at 3. Adding concrete trigger phrases would lift the weakest dimension.

Suggestions

Append an explicit trigger clause, e.g. "Use when setting up or running end-to-end tests, automating browser test suites, or adding E2E checks to CI/CD pipelines."

Add natural synonyms users say, such as "E2E tests", "automated browser tests", and "Playwright tests", to broaden trigger coverage.

Tighten capabilities from labels to concrete operations (e.g., "record and replay test scripts, capture screenshots and traces") to push specificity toward 5.

DimensionReasoningScore

Specificity

The description lists several concrete capability areas — "browser automation, visual regression, cross-browser testing, and CI/CD integration" — alongside the named tool (Playwright), which matches the anchor listing several specific actions with minor gaps. It is not a 5 because the actions are capability labels rather than granular concrete operations, and not a 3 because more than 1-2 specific actions are named.

4 / 5

Completeness

The description clearly states what the skill does ("End-to-end testing workflow with Playwright for browser automation, visual regression, cross-browser testing, and CI/CD integration") but provides no "Use when..." clause or equivalent explicit trigger guidance, so per the judging guideline completeness is capped at 3. It is not a 2 because the 'what' is clear and specific, not vague.

3 / 5

Trigger Term Quality

It includes natural terms a user would say — "End-to-end testing", "Playwright", "browser automation", "visual regression", "cross-browser testing", "CI/CD integration" — giving good keyword coverage. It is not a 5 because common synonyms/variations (e.g., "E2E tests", "automated testing", "Playwright tests") are absent, and not a 3 because coverage goes well beyond a single relevant keyword.

4 / 5

Distinctiveness Conflict Risk

Naming Playwright plus the four scoped testing capabilities carves a mostly-distinct niche with only minor overlap risk against closely related general testing skills. It is not a 5 because "end-to-end testing" alone is a broad category that could overlap with other testing skills, and not a 3 because the tool-specific framing meaningfully narrows the trigger space.

4 / 5

Total

15

/

20

Passed

Validation

93%

Checks the skill against the spec for correct structure and formatting. All validation checks must pass before discovery and implementation can be scored.

Validation15 / 16 Passed

Validation for skill structure

CriteriaDescriptionResult

frontmatter_unknown_keys

Unknown frontmatter key(s) found; consider removing or moving to metadata

Warning

Total

15

/

16

Passed

Repository
administrakt0r/AI-Agents-Safe-Coding-Skills
Reviewed

Table of Contents

Is this your skill?

If you maintain this skill, you can claim it as your own. Once claimed, you can manage eval scenarios, bundle related skills, attach documentation or rules, and ensure cross-agent compatibility.