CtrlK
BlogDocsLog inGet started
Tessl Logo

playwright-core

Battle-tested Playwright patterns for E2E, API, component, visual, accessibility, and security testing. Covers locators, assertions, fixtures, network mocking, auth flows, debugging, and framework recipes for React, Next.js, Vue, and Angular. TypeScript and JavaScript.

62

Quality

74%

Does it follow best practices?

Run evals on this skill

Adds up to 20 points to the overall score

View guide

SecuritybySnyk

Passed

No findings from the security scan

Fix and improve this skill with Tessl

tessl review fix ./skills/playwright/core/SKILL.md
SKILL.md
Quality
Evals
Security

Quality

Content

82%Weight 40%Scale 1-5

Reviews the quality of instructions and guidance provided to agents. Good implementation is clear, handles edge cases, and produces reliable results.

An excellent progressive-disclosure design: a lean, opinionated overview with ten concrete golden rules and task-oriented routing tables, with no filler. The main deductions are the absence of any inline executable snippet and the fact that the referenced guide files are not present in the bundle, leaving the index unverified.

DimensionReasoningScore

Conciseness

The body is a lean index plus ten terse, dense golden rules ("Never `page.waitForTimeout()`", "Retries: `2` in CI, `0` locally") with zero explanation of concepts Claude already knows. Every token earns its place; the only near-redundancy (the tagline and the "46 reference guides" sentence) is trivial.

5 / 5

Actionability

The golden rules are highly concrete — exact APIs (`getByRole()`, `expect(locator).toBeVisible()`, `test.extend()`), exact config values (retries `2`/`0`, traces `'on-first-retry'`, `baseURL` in config) — and the routing tables give unambiguous lookup paths. It stops short of 5 because the body contains no actual executable code or config snippet, only directives.

4 / 5

Workflow Clarity

The skill is an index/router, and its single action — match your task or problem to a guide — is unambiguous via the "What you're doing → Guide" and "Problem → Guide" tables, and the "Architecture Decisions" table covers decision points. It misses 5 because no explicit sequenced workflow or checkpoints (e.g., how to combine guides when debugging a failure) are described in the body.

4 / 5

Progressive Disclosure

Structurally exemplary: SKILL.md is a pure overview with one-level-deep, well-categorized references organized by task, problem, framework, and architecture question. However, the linked files (locators.md, debugging.md, etc.) and the "46 reference guides" count cannot be verified — no references/ or other bundle directory exists alongside the skill — so navigation cannot be confirmed as working.

4 / 5

Total

17

/

20

Passed

Description

66%Weight 40%Scale 1-5

Based on the skill's description, can an agent find and select it at the right time? Clear, specific descriptions lead to better discovery.

A strong, information-dense description that clearly conveys breadth of capability and names the exact tool and frameworks. Its single material weakness is the complete absence of a "Use when..." trigger clause, which both caps completeness and slightly weakens its usefulness for skill selection.

Suggestions

Append an explicit trigger clause, e.g., "Use when writing, debugging, or fixing Playwright tests, or when the user mentions Playwright, E2E tests, browser automation, or flaky tests."

Drop the puffery word "Battle-tested" — it adds no information a trigger-matching model or user can act on.

Add the natural synonyms users actually say — "end-to-end", "browser automation", "flaky tests" — to improve trigger-term coverage.

DimensionReasoningScore

Specificity

The description lists several concrete capability areas — "locators, assertions, fixtures, network mocking, auth flows, debugging" plus framework recipes for four named frameworks — matching the anchor for several specific actions with minor gaps. It falls short of 5 because "Battle-tested Playwright patterns" is mild puffery and it names topics rather than concrete actions; it is well above 3 because coverage is broad and domain-specific, not just 1-2 actions.

4 / 5

Completeness

The 'what' is clear and comprehensive (testing types, techniques, and frameworks covered), but there is no "Use when..." clause or any equivalent explicit trigger guidance anywhere in the description. Per the rubric guideline, a missing 'Use when...' clause caps completeness at 3; it is not lower because the 'what' half is strong.

3 / 5

Trigger Term Quality

Good keyword coverage with natural terms users would say: "Playwright", "E2E", "API testing", "visual", "accessibility", "debugging", "React, Next.js, Vue, and Angular", "TypeScript and JavaScript". A few common natural phrases are missing (e.g., "end-to-end", "browser automation", "flaky tests", "test runner"), which keeps it below the comprehensive-synonym anchor of 5 but clearly above 3.

4 / 5

Distinctiveness Conflict Risk

Naming Playwright explicitly gives it a clear niche distinct from generic testing skills, but the broad sweep across E2E, API, component, visual, accessibility, and security testing creates minor overlap risk with dedicated testing or security skills. It is more distinct than the anchor 3 example but lacks the explicit trigger scoping that would justify 5.

4 / 5

Total

15

/

20

Passed

Validation

93%

Checks the skill against the spec for correct structure and formatting. All validation checks must pass before discovery and implementation can be scored.

Validation — 15 / 16 Passed

Validation for skill structure

CriteriaDescriptionResult

relative_links

Relative link issues: 49 missing

Warning

Total

15

/

16

Passed

Repository
zebbern/claude-code-guide
Reviewed

Table of Contents

Is this your skill?

If you maintain this skill, you can claim it as your own. Once claimed, you can manage eval scenarios, bundle related skills, attach documentation or rules, and ensure cross-agent compatibility.