CtrlK
BlogDocsLog inGet started
Tessl Logo

playwright-cli

Automates browser interactions for web testing, form filling, screenshots, and data extraction. Use when the user needs to navigate websites, interact with web pages, fill forms, take screenshots, test web applications, or extract information from web pages.

89

22.50x
Quality

86%

Does it follow best practices?

Impact

90%

22.50x

Average score across 3 eval scenarios

SecuritybySnyk

High

Do not use without reviewing

SKILL.md
Quality
Evals
Security

Quality

Content

90%Weight 40%Scale 1-5

Reviews the quality of instructions and guidance provided to agents. Good implementation is clear, handles edge cases, and produces reliable results.

An exemplary lean, command-first reference: fully executable commands, no filler, and worked examples for the main use cases. The only room for improvement is explicit error-recovery guidance in the workflow and potentially splitting the command catalog into a reference file.

DimensionReasoningScore

Conciseness

The body is almost entirely executable command listings with zero padded prose and no explanation of concepts Claude already knows — every token earns its place, matching the lean anchor.

5 / 5

Actionability

Commands like "playwright-cli fill e5 \"user@example.com\"" are fully executable and copy-paste ready, and the worked examples (form submission, multi-tab workflow, debugging) cover the common cases per the anchor-5 example.

5 / 5

Workflow Clarity

The core workflow ("1. Navigate... 2. Interact using refs from the snapshot 3. Re-snapshot after significant changes") is a clear sequence with the re-snapshot acting as a checkpoint, but there is no explicit verification or error-recovery guidance (e.g., what to do when a ref is stale), so it matches anchor 4 rather than the validation-rich anchor 5.

4 / 5

Progressive Disclosure

A single, well-sectioned command reference (Core/Navigation/Keyboard/Mouse/Tabs/DevTools/Sessions) with no nested references is appropriately organized for a small skill, though the full command catalog could be split into a references/ file with SKILL.md keeping quick-start and workflow, keeping it below the anchor-5 split pattern.

4 / 5

Total

18

/

20

Passed

Description

83%Weight 40%Scale 1-5

Based on the skill's description, can an agent find and select it at the right time? Clear, specific descriptions lead to better discovery.

A strong description that clearly states both capabilities and explicit usage triggers in third person with concrete action verbs. The main gaps are missing common synonyms (scraping, automation, playwright) and omission of several CLI capabilities, which leave it just short of top marks on specificity and trigger quality.

DimensionReasoningScore

Specificity

Lists several concrete actions ("web testing, form filling, screenshots, and data extraction") but omits capabilities the CLI supports such as console/network debugging, tab management, and PDF export, matching the 'minor gaps in coverage' anchor rather than the comprehensive anchor 5.

4 / 5

Completeness

Explicitly answers both what ("Automates browser interactions for web testing, form filling, screenshots, and data extraction") and when ("Use when the user needs to navigate websites...") with concrete trigger phrases, directly matching the anchor-5 example pattern.

5 / 5

Trigger Term Quality

Good natural keyword coverage ("navigate websites", "fill forms", "take screenshots", "test web applications", "extract information from web pages"), but missing common variations like "scrape/web scraping", "browser automation", or "playwright", so it falls short of the comprehensive-synonyms anchor.

4 / 5

Distinctiveness Conflict Risk

"Automates browser interactions" carves a clear niche, but "take screenshots" and "extract information from web pages" create minor overlap risk with dedicated screenshot and web-scraping skills, fitting the 'mostly distinct; minor overlap risk' anchor.

4 / 5

Total

17

/

20

Passed

Validation

100%

Checks the skill against the spec for correct structure and formatting. All validation checks must pass before discovery and implementation can be scored.

Validation — 16 / 16 Passed

Validation for skill structure

No warnings or errors.

Repository
Dicklesworthstone/pi_agent_rust
Reviewed

Table of Contents

Is this your skill?

If you maintain this skill, you can claim it as your own. Once claimed, you can manage eval scenarios, bundle related skills, attach documentation or rules, and ensure cross-agent compatibility.