CtrlK
BlogDocsLog inGet started
Tessl Logo

control-ui

Build or adapt a local browser/CDP harness to drive and inspect a web, IDE, or Electron UI. Use for local UI verification, screenshots, accessibility snapshots, perf profiles, visual diffs, or reproducing UI bugs.

68

Quality

81%

Does it follow best practices?

Run evals on this skill

Adds up to 20 points to the overall score

View guide

SecuritybySnyk

Passed

No findings from the security scan

SKILL.md
Quality
Evals
Security

Quality

Content

75%Weight 40%Scale 1-5

Reviews the quality of instructions and guidance provided to agents. Good implementation is clear, handles edge cases, and produces reliable results.

A tight, well-structured operational skill: executable probe code, a disciplined interaction loop with capture/verify steps, and thoughtful guardrails against selector staleness and privacy leaks. Its weaknesses are mild — a small redundancy with the description, no failure path after verification, and a body that could shed ~40 lines by moving the harness code into bundle files.

Suggestions

Add a failure branch to the Interaction Loop, e.g. "If the expected state change did not occur, re-capture and diff against the before snapshot before retrying or reporting."

Trim the "What It Is Used For" section, which largely repeats the frontmatter description, or fold its one new item (before/after evidence for verify-this) into the opening paragraph.

Consider moving the two full harness code blocks into a scripts/ or references/ file and keeping only a one-line invocation in SKILL.md to reduce always-loaded tokens.

DimensionReasoningScore

Conciseness

The body is lean — terse bullets, two short code probes, no explanations of concepts Claude already knows (no "what is CDP" primer). Minor trimmable fat remains: the "What It Is Used For" list partially restates the frontmatter description's use cases. Not score 5 because of that redundancy; not score 3 because almost every token carries operational guidance.

4 / 5

Actionability

Two near-executable Playwright probes (web and CDP) plus concrete steps ("Select the correct page by stable app markers, not by tab order alone", "Prefer accessibility roles, labels, and stable data-* selectors") give directly usable guidance. Not score 5 because the code contains repo-dependent placeholders ("<port>", "<app-root-selector>") and, appropriately flagged, is not literally copy-paste ready; not score 3 because the placeholders are explicitly explained and the examples cover the common web and Electron cases.

4 / 5

Workflow Clarity

The Setup Pattern (6 ordered steps) and Interaction Loop (snapshot → one action → snapshot → verify expected state) are clearly sequenced, and page selection has an error-recovery fallback ("If no page matches, list available page titles and URLs instead of guessing" plus the thrown error in code). Not score 5 because the interaction loop's "Verify the expected state change" step has no explicit failure path (what to do when verification fails); not score 3 because checkpoints are largely explicit and the operations are not destructive or batch.

4 / 5

Progressive Disclosure

No bundle files exist, and the self-contained body is well organized into clearly labeled sections (Setup Pattern, Generic Web Harness, Generic CDP Harness, Interaction Loop, CDP Capabilities, Page Selection, Guardrails) with no nested or buried references. Not score 5 because the body runs ~105 lines — past the point where the short self-contained exception applies — and the two harness probes or the CDP Capabilities list could live in a scripts/ or references/ file; not score 3 because everything inline is compact, appropriately placed, and easy to navigate.

4 / 5

Total

16

/

20

Passed

Description

88%Weight 40%Scale 1-5

Based on the skill's description, can an agent find and select it at the right time? Clear, specific descriptions lead to better discovery.

A strong description: it states a concrete capability set in third person, includes an explicit "Use for..." trigger clause, and covers the domain comprehensively. The main improvement is adding a few more natural synonyms (e.g., "browser automation", tool names like Playwright) that a user might actually say.

Suggestions

Add one or two natural synonyms users would say, such as "browser automation" or "drive the app with Playwright", to broaden trigger coverage.

Sharpen the when-clause to reduce overlap with generic verification skills, e.g., "Use when a UI change needs real-browser evidence (screenshots, a11y snapshots, perf traces)".

DimensionReasoningScore

Specificity

Phrases like "Build or adapt a local browser/CDP harness", "drive and inspect a web, IDE, or Electron UI", "screenshots, accessibility snapshots, perf profiles, visual diffs, or reproducing UI bugs" enumerate multiple concrete, distinct capabilities with comprehensive coverage of the skill's domain. Not score 4 because there is no meaningful coverage gap; not below because every listed action is specific rather than generic.

5 / 5

Completeness

The "what" is explicit ("Build or adapt a local browser/CDP harness to drive and inspect a web, IDE, or Electron UI") and the "when" is explicit with concrete triggers ("Use for local UI verification, screenshots, accessibility snapshots, perf profiles, visual diffs, or reproducing UI bugs"). Not score 4 because the when-clause is already specific and trigger-phrase style, matching the anchor-5 example's structure.

5 / 5

Trigger Term Quality

Natural user phrases are present ("UI verification", "screenshots", "visual diffs", "reproducing UI bugs", "Electron"), but common synonyms users would plausibly say — "browser automation", "take a screenshot of the app", "Playwright/Puppeteer" — are absent. Not score 5 because synonym/extension coverage is incomplete; not score 3 because the included terms are varied and genuinely match what a user would say, not just technical jargon.

4 / 5

Distinctiveness Conflict Risk

The local browser/CDP harness niche is clear and the trigger list is fairly distinct, but terms like "local UI verification" and "screenshots" have minor overlap risk with generic verify/run or screenshot skills. Not score 5 because that overlap is non-trivial; not score 3 because the CDP/Electron/accessibility-snapshot framing is well outside what a generic testing skill would claim.

4 / 5

Total

18

/

20

Passed

Validation

100%

Checks the skill against the spec for correct structure and formatting. All validation checks must pass before discovery and implementation can be scored.

Validation — 16 / 16 Passed

Validation for skill structure

No warnings or errors.

Repository
cursor/plugins
Reviewed

Table of Contents

Is this your skill?

If you maintain this skill, you can claim it as your own. Once claimed, you can manage eval scenarios, bundle related skills, attach documentation or rules, and ensure cross-agent compatibility.