CtrlK
BlogDocsLog inGet started
Tessl Logo

playwright-cli

官方Microsoft Playwright CLI网页自动化工具,支持所有主流浏览器的无头/有头自动化操作,包括页面导航、元素交互、截图、录制、测试等功能。当用户提到网页自动化、浏览器操作、爬虫、截图、录制用户操作、E2E测试时触发。

66

Quality

80%

Does it follow best practices?

Run evals on this skill

Adds up to 20 points to the overall score

View guide

SecuritybySnyk

Critical

Do not install without reviewing

Fix and improve this skill with Tessl

tessl review fix ./configs/microservice/bff-service/configs/agent-skills/clawhub/playwright-cli-openclaw/SKILL.md
SKILL.md
Quality
Evals
Security

Quality

Content

76%Weight 40%Scale 1-5

Reviews the quality of instructions and guidance provided to agents. Good implementation is clear, handles edge cases, and produces reliable results.

A highly actionable, well-organized command reference: nearly every line is an executable command or complete script, and sections are clearly headed. The main gaps are the total absence of validation/verification steps — most visibly in the batch-screenshot example, which caps workflow clarity — and modest opportunities to tighten the intro and split the Python API example into a reference file.

Suggestions

Add a verification step to the batch screenshot example (示例1), e.g. checking each output file exists and is non-empty after the loop, which would lift workflow clarity past the batch-operation cap.

Trim the opening paragraph, which repeats the frontmatter description almost verbatim and adds no new information.

Move the full Python script (示例2) into a references/ file (e.g. references/examples.md) and keep only a one-line pointer in SKILL.md, tightening the main body.

DimensionReasoningScore

Conciseness

The body is dominated by lean, commented command blocks with essentially no explanation of concepts Claude already knows. Minor trimming is possible — the opening line "Playwright CLI是微软官方的浏览器自动化工具...提供强大的网页自动化能力" restates the frontmatter description, and a few comments ("# 或者", "# 安装系统依赖(Linux环境)") are slightly redundant — so it sits at the 4 anchor ("efficient; minor instances of over-explanation") rather than the 5 anchor's every-token-earns-its-place.

4 / 5

Actionability

Nearly everything is copy-paste ready: exact commands with flags for screenshots ("playwright screenshot https://example.com --full-page full.png"), codegen targets, PDF generation, and test runs, plus a complete, executable Python script in 示例2 and concrete device-emulation commands in 示例3. This matches the 5 anchor — fully executable commands covering the common cases — and is clearly above the 4 anchor, which allows minor gaps.

5 / 5

Workflow Clarity

The install sequence (CLI → browsers → system deps) and section ordering are clear, but there are no validation or verification checkpoints anywhere. Critically, 示例1 is a batch operation — a loop screenshotting multiple URLs — with no step verifying the output files were created, which per the rubric caps workflow_clarity at 3 even though the sequence itself would otherwise merit 4.

3 / 5

Progressive Disclosure

No bundle files exist (no references/, scripts/, or assets/ directories), so the skill is a single self-contained document; it is scored on its internal structure. Sections are well-organized and clearly headed (安装步骤, 核心常用命令, 使用示例, 最佳实践) with nothing buried. It stays at 4 rather than 5 because at ~125 lines the full Python API example (示例2) is content that could live in a separate reference file, and no external reference structure exists at all.

4 / 5

Total

16

/

20

Passed

Description

83%Weight 40%Scale 1-5

Based on the skill's description, can an agent find and select it at the right time? Clear, specific descriptions lead to better discovery.

A strong description that pairs an explicit, concrete capability list with an explicit trigger clause in the "当用户提到...时触发" (triggers when the user mentions...) form. Trigger terms are natural user phrasings. The only weaknesses are the hedged "等功能" (etc.) tail on the capability list and a couple of broad trigger terms (crawling, screenshot) with mild conflict risk.

DimensionReasoningScore

Specificity

The description lists several concrete actions — "页面导航、元素交互、截图、录制、测试" (page navigation, element interaction, screenshots, recording, testing) — plus a concrete modality distinction ("无头/有头自动化操作", headless/headed). It falls short of the 5 anchor because "等功能" ("etc. functions") is hedging that leaves coverage gaps rather than enumerating comprehensively, but it clearly exceeds the 3 anchor's "1-2 concrete actions".

4 / 5

Completeness

It explicitly answers both questions: the "what" is a list of concrete capabilities (navigation, interaction, screenshots, recording, testing), and the "when" is an explicit trigger clause — "当用户提到网页自动化、浏览器操作、爬虫、截图、录制用户操作、E2E测试时触发" — which is the direct equivalent of a "Use when..." clause with concrete trigger phrases. This matches the 5 anchor exactly and exceeds the 4 anchor, whose 'when' is less explicit.

5 / 5

Trigger Term Quality

Trigger terms are natural phrases users would actually say: "网页自动化" (web automation), "浏览器操作" (browser operation), "爬虫" (crawling), "截图" (screenshot), "录制用户操作" (recording user actions), "E2E测试". A few common variations are missing (e.g., "自动化测试", "模拟浏览器", the literal "Playwright" as a trigger phrase), keeping it below the 5 anchor's comprehensive synonym/extension coverage, but it is well above the 3 anchor's partial coverage.

4 / 5

Distinctiveness Conflict Risk

The niche is mostly distinct: "官方Microsoft Playwright CLI网页自动化工具" names a specific official tool, and most triggers are web-automation-specific. It stays at 4 rather than 5 because broad triggers like "爬虫" (crawling) and "截图" (screenshot) could plausibly fire for other scraping or screenshot skills, creating minor overlap risk; it is clearly better than the 3 anchor's generic overlap.

4 / 5

Total

17

/

20

Passed

Validation

100%

Checks the skill against the spec for correct structure and formatting. All validation checks must pass before discovery and implementation can be scored.

Validation — 16 / 16 Passed

Validation for skill structure

No warnings or errors.

Repository
UnicomAI/wanwu
Reviewed

Table of Contents

Is this your skill?

If you maintain this skill, you can claim it as your own. Once claimed, you can manage eval scenarios, bundle related skills, attach documentation or rules, and ensure cross-agent compatibility.