CtrlK
BlogDocsLog inGet started
Tessl Logo

playwright-app-testing

Test the Expensify App using Playwright browser automation. Use when user requests browser testing, after making frontend changes, or when debugging UI issues

66

Quality

80%

Does it follow best practices?

Run evals on this skill

Adds up to 20 points to the overall score

View guide

SecuritybySnyk

High

Do not use without reviewing

SKILL.md
Quality
Evals
Security

Quality

Content

78%Weight 40%Scale 1-5

Reviews the quality of instructions and guidance provided to agents. Good implementation is clear, handles edge cases, and produces reliable results.

A well-organized, actionable testing skill with concrete commands and a clear workflow. The main improvements are tightening the redundant use-case examples and adding a concrete Playwright interaction snippet.

Suggestions

Condense or remove the 'Example Usage' scenarios since they restate the 'When to Use' section, recovering tokens without losing clarity.

Add a short concrete Playwright MCP snippet (e.g., a snapshot or click command) to make the 'Interact' step executable rather than descriptive.

DimensionReasoningScore

Conciseness

The body is largely efficient with concrete commands and minimal explanation of known concepts, but the 'When to Use' section partially restates the description and the three-scenario Example Usage block is somewhat padded.

4 / 5

Actionability

It provides executable commands (server check, sed edits, specific URL and credentials) but the Playwright interaction step ('inspect, click, type, and navigate') is described abstractly without concrete code, leaving minor gaps.

4 / 5

Workflow Clarity

A clear sequenced workflow (verify server -> navigate -> interact) with a validation checkpoint for the server and explicit guidance on snapshotting instead of arbitrary waits, though it lacks an explicit error-recovery feedback loop.

4 / 5

Progressive Disclosure

The skill is self-contained with no bundle files and is organized into clear, well-signaled sections (When to Use, Prerequisites, Workflow, Sign-In, When NOT to Use) appropriate for a single-purpose skill that needs no external references.

5 / 5

Total

17

/

20

Passed

Description

82%Weight 40%Scale 1-5

Based on the skill's description, can an agent find and select it at the right time? Clear, specific descriptions lead to better discovery.

A strong, well-structured description that clearly states both capability and trigger conditions with natural phrasing. It is slightly light on the breadth of concrete actions and a few natural synonyms.

Suggestions

Expand the 'what' clause with 1-2 more concrete actions (e.g., 'inspect elements, click, type, and assert UI behavior') to lift specificity toward comprehensive coverage.

Add common synonyms like 'E2E', 'end-to-end testing', or 'Playwright' as explicit trigger terms so the description matches more natural user phrasings.

DimensionReasoningScore

Specificity

The description names the domain ('Playwright browser automation') and one concrete action ('Test the Expensify App') but does not enumerate multiple specific actions, matching the anchor for 1-2 concrete actions without comprehensive coverage.

3 / 5

Completeness

It explicitly answers both 'what' ('Test the Expensify App using Playwright browser automation') and 'when' with concrete trigger phrases ('Use when user requests browser testing, after making frontend changes, or when debugging UI issues'), matching the top anchor.

5 / 5

Trigger Term Quality

It includes natural phrases users would say ('browser testing', 'frontend changes', 'debugging UI issues') but omits common synonyms like 'E2E', 'end-to-end', or 'Playwright' as a trigger term, fitting the good-but-incomplete anchor.

4 / 5

Distinctiveness Conflict Risk

The combination of 'Expensify App' and 'Playwright browser automation' carves a clear niche with distinct triggers and minimal overlap with other skills.

5 / 5

Total

17

/

20

Passed

Validation

93%

Checks the skill against the spec for correct structure and formatting. All validation checks must pass before discovery and implementation can be scored.

Validation15 / 16 Passed

Validation for skill structure

CriteriaDescriptionResult

frontmatter_unknown_keys

Unknown frontmatter key(s) found; consider removing or moving to metadata

Warning

Total

15

/

16

Passed

Repository
Expensify/App
Reviewed

Table of Contents

Is this your skill?

If you maintain this skill, you can claim it as your own. Once claimed, you can manage eval scenarios, bundle related skills, attach documentation or rules, and ensure cross-agent compatibility.