CtrlK
BlogDocsLog inGet started
Tessl Logo

playwright-testing

Playwright E2E testing patterns, cross-browser configuration, page objects, and CI setup. Use when creating E2E specs, visual regression suites, or configuring Playwright in CI. Trigger terms: playwright, e2e, trace, page object, cross-browser

84

1.16x
Quality

88%

Does it follow best practices?

Impact

90%

1.16x

Average score across 2 eval scenarios

SecuritybySnyk

Passed

No findings from the security scan

SKILL.md
Quality
Evals
Security

Quality

Content

93%Weight 40%Scale 1-5

Reviews the quality of instructions and guidance provided to agents. Good implementation is clear, handles edge cases, and produces reliable results.

An exemplary reference-style skill body: token-efficient, assumes competence, and packed with executable commands, exact config values, and hard-won gotchas (parallel-worker state leaks, forbidOnly in CI). The only structural limitation is that, as a topic reference rather than a workflow, it has no explicit step sequence with validation checkpoints — appropriate for this skill type.

DimensionReasoningScore

Conciseness

The body is lean and dense: every line carries non-obvious information (locator priority chain, CI-specific config values like "retries: 2 and workers: 1 in CI trade wall-clock for determinism") and assumes Playwright competence. No padded explanation of what E2E testing is — matches the anchor 5 'every token earns its place' example.

5 / 5

Actionability

The commands block is copy-paste ready ("npx playwright test --ui", "npx playwright codegen http://localhost:3000"), config guidance uses exact values, and the mock pattern is real code ("await page.route('/api/login', route => route.fulfill({ status: 200, body: ... }))"). The single ellipsis is trivial; this fits anchor 5's fully-executable coverage of common cases.

5 / 5

Workflow Clarity

This is a patterns/reference skill rather than a sequential process, so there is no multi-step workflow to sequence; verification guidance is present but implicit ("show-report # HTML report (verify 0 failures)", "--debug # step through", attaching trace IDs to failure tickets). Fits anchor 4's 'clear with most checkpoints' better than 5's explicit validate-fix-retry loops, which this format doesn't need given no destructive/batch operations.

4 / 5

Progressive Disclosure

Under 50 lines with no bundle files, organized into clean sections (Layout, Locator priority, Gotchas, Commands), and the only pointers are one level deep and clearly signaled (project config at .opencastle/stack/testing-config.md and the official docs URL). Per the rubric's simple-skill guideline, this scores 5 with just well-organized sections.

5 / 5

Total

19

/

20

Passed

Description

83%Weight 40%Scale 1-5

Based on the skill's description, can an agent find and select it at the right time? Clear, specific descriptions lead to better discovery.

A strong description with a clear what-and-when structure and a dedicated trigger-terms clause. It is fully complete and specific; the only weaknesses are a few missing natural trigger synonyms and a couple of generic trigger terms ("e2e", "trace") that create minor overlap risk with adjacent skills.

Suggestions

Add missing natural trigger synonyms such as "end-to-end", "browser test", "flaky", and "screenshot" to the trigger-terms list.

Qualify the generic term "trace" (e.g., "Playwright trace") so it cannot fire for logging/tracing or other E2E-framework skills.

Consider naming one or two concrete actions in the opening sentence (e.g., "Write and debug E2E specs...") to lift specificity from domain naming to explicit actions.

DimensionReasoningScore

Specificity

"Playwright E2E testing patterns, cross-browser configuration, page objects, and CI setup" lists several specific capability areas with only minor gaps. Falls just short of anchor 5 because "testing patterns" is domain naming rather than the multiple concrete actions the top anchor exemplifies.

4 / 5

Completeness

The description explicitly answers both: what ("E2E testing patterns, cross-browser configuration, page objects, and CI setup") and when ("Use when creating E2E specs, visual regression suites, or configuring Playwright in CI"), plus a dedicated trigger-terms clause. This matches the anchor 5 example structure exactly.

5 / 5

Trigger Term Quality

"playwright, e2e, trace, page object, cross-browser" are natural terms users would say, but common variations are missing: "end-to-end", "browser test", "flaky", "screenshot", and "visual regression" appears only in the when-clause. Fits anchor 4 (good coverage, a few natural terms missing) rather than 5's comprehensive synonym coverage.

4 / 5

Distinctiveness Conflict Risk

The Playwright niche is clear, but "e2e" and especially "trace" are generic terms that could overlap with other E2E frameworks (Cypress, Selenium) or logging/tracing skills. Fits anchor 4 (mostly distinct, minor overlap risk) better than 5's minimal conflict risk.

4 / 5

Total

17

/

20

Passed

Validation

100%

Checks the skill against the spec for correct structure and formatting. All validation checks must pass before discovery and implementation can be scored.

Validation — 16 / 16 Passed

Validation for skill structure

No warnings or errors.

Repository
monkilabs/opencastle
Reviewed

Table of Contents

Is this your skill?

If you maintain this skill, you can claim it as your own. Once claimed, you can manage eval scenarios, bundle related skills, attach documentation or rules, and ensure cross-agent compatibility.