CtrlK
BlogDocsLog inGet started
Tessl Logo

agent-browser

Browser automation CLI for AI agents. Use when the user needs to inspect, test, or automate browser behavior: navigating pages, filling forms, clicking buttons, taking screenshots, extracting page data, reading selected OpenDesign browser-tab context, testing web apps, dogfooding OpenDesign previews, QA, bug hunts, or reviewing app quality. Prefer local OpenDesign preview URLs unless the user explicitly asks for external browsing.

72

Quality

88%

Does it follow best practices?

Run evals on this skill

Adds up to 20 points to the overall score

View guide

SecuritybySnyk

Passed

No findings from the security scan

SKILL.md
Quality
Evals
Security

Quality

Content

81%Weight 40%Scale 1-5

Reviews the quality of instructions and guidance provided to agents. Good implementation is clear, handles edge cases, and produces reliable results.

A dense, operational skill body with excellent error handling, explicit expected outcomes, and strong safety rules. Its weaknesses are moderate: a verbatim-duplicated CDP startup block, and core click/type syntax being deferred entirely to a runtime-fetched upstream guide, which costs it both actionability and conciseness.

Suggestions

Consolidate the duplicated CDP launch/polling script: define the 'ensure CDP is up' block once (e.g., a single copy in the CDP Startup Contract section) and reference it from the Smoke Path instead of repeating ~20 lines verbatim.

Inline the handful of core interaction commands (click, type, wait, element screenshot syntax) in a short reference table so the most common path does not require shelling out to `agent-browser skills get core` before every interaction.

Replace the hedged 'when the core guide exposes an element-screenshot command' with the actual command name and arguments once verified, making element screenshots copy-paste ready.

DimensionReasoningScore

Conciseness

The body is lean and command-first — it assumes competence, explains no general browser concepts, and every section earns its tokens. It is not a 5 because the full CDP launch/polling block (~20 lines) is duplicated verbatim in 'CDP Startup Contract' and 'OpenDesign Smoke Path', and the Chrome manual-launch command also appears twice, which could be consolidated.

4 / 5

Actionability

Most guidance is copy-paste ready: install check, connect sequence with curl polling, snapshot/screenshot commands, cleanup trap, and expected success criteria. It is not a 5 because core interaction commands (click, type, wait) are never shown — they are deferred entirely to `agent-browser skills get core` — and the element screenshot is hedged as 'when the core guide exposes an element-screenshot command', leaving the common click/type path without concrete syntax inline.

4 / 5

Workflow Clarity

The 11-step Workflow is clearly sequenced with explicit checkpoints: verify install, confirm CDP before connecting, snapshot before selecting elements, re-snapshot after state changes, and report title/URL/text/screenshot. Error recovery is specified for each failure mode (Chrome crash message, manual-launch fallback, stop if CLI missing), and the smoke path defines expected success outputs (title 'OpenDesign', URL under 127.0.0.1:17573, screenshot path). This matches the anchor-5 feedback-loop pattern; a 4 would lack the fallback and expected-result definitions.

5 / 5

Progressive Disclosure

Structure is good: well-labeled sections, and heavy upstream material (electron, slack, dogfood, sandbox, agentcore guides) is loaded on demand via `agent-browser skills get` and redirected to temp files — a genuine one-level-deep disclosure pattern. It is not a 5 because the core command reference is only reachable through a runtime CLI call rather than a clearly signaled in-bundle reference, and the duplicated CDP block suggests content that belongs in one place, leaving minor organization gaps.

4 / 5

Total

17

/

20

Passed

Description

96%Weight 40%Scale 1-5

Based on the skill's description, can an agent find and select it at the right time? Clear, specific descriptions lead to better discovery.

A strong description: it states what the CLI does and when to use it in third person, with concrete trigger phrases and good natural-language keyword coverage. The only weakness is a broad generic opening that slightly blurs its niche versus general browser-automation skills.

DimensionReasoningScore

Specificity

The description lists multiple concrete actions — 'navigating pages, filling forms, clicking buttons, taking screenshots, extracting page data, reading selected OpenDesign browser-tab context, testing web apps, dogfooding' — giving comprehensive coverage of the CLI's capabilities with no padding. It is not a 4 because coverage spans inspection, interaction, extraction, and QA use cases rather than leaving minor gaps.

5 / 5

Completeness

Both halves are explicit: 'Browser automation CLI for AI agents' answers what, and 'Use when the user needs to inspect, test, or automate browser behavior: ...' answers when with concrete trigger phrases. This matches the anchor-5 good example structure exactly, so it is not a 4.

5 / 5

Trigger Term Quality

Natural user phrases are well covered within the description itself: 'take screenshots', 'filling forms', 'clicking buttons', 'QA', 'bug hunts', 'dogfooding', 'testing web apps', 'reviewing app quality' — matching how a user would actually phrase these needs. It is not a 4 because both common synonyms (QA/bug hunt, inspect/test) and informal variants (dogfooding) are present.

5 / 5

Distinctiveness Conflict Risk

The OpenDesign scoping ('reading selected OpenDesign browser-tab context, dogfooding OpenDesign previews, Prefer local OpenDesign preview URLs') gives it a clear niche, but the opening 'Browser automation CLI for AI agents' plus generic capabilities ('extracting page data', 'QA') leaves minor overlap risk with general browser-automation or scraping skills. It does not reach 5 because the capability envelope is broad enough that a generic 'automate the browser' request could plausibly match competing skills.

4 / 5

Total

19

/

20

Passed

Validation

93%

Checks the skill against the spec for correct structure and formatting. All validation checks must pass before discovery and implementation can be scored.

Validation — 15 / 16 Passed

Validation for skill structure

CriteriaDescriptionResult

frontmatter_unknown_keys

Unknown frontmatter key(s) found; consider removing or moving to metadata

Warning

Total

15

/

16

Passed

Repository
nexu-io/open-design
Reviewed

Table of Contents

Is this your skill?

If you maintain this skill, you can claim it as your own. Once claimed, you can manage eval scenarios, bundle related skills, attach documentation or rules, and ensure cross-agent compatibility.