CtrlK
BlogDocsLog inGet started
Tessl Logo

browser-use

Control the user's existing signed-in Chrome: callable Codex Chrome plugin first, then OpenClaw extension-backed mcporter, with direct DevTools attachment only as a last fallback.

57

Quality

65%

Does it follow best practices?

Run evals on this skill

Adds up to 20 points to the overall score

View guide

SecuritybySnyk

Low

Low-risk findings worth noting

Fix and improve this skill with Tessl

tessl review fix ./skills/browser-use/SKILL.md
SKILL.md
Quality
Evals
Security

Quality

Content

77%Weight 40%Scale 1-5

Reviews the quality of instructions and guidance provided to agents. Good implementation is clear, handles edge cases, and produces reliable results.

The body is an unusually strong operational skill: executable commands, a fail-closed readiness checklist, and explicit feedback loops for nearly every failure mode. Its weaknesses are structural — a monolithic ~290-line file with no reference split, repeated command prefixes, and inline version-specific notes that inflate token cost without adding clarity.

Suggestions

Split the Discovery Timeout/Daemon Configuration and Extension Setup and Repair sections into reference files (e.g. references/daemon-config.md, references/setup.md) and keep SKILL.md as a route + readiness overview with one-level-deep pointers.

State the 'MCPORTER_CHROME_DEVTOOLS_RELAY_POLICY=require' prefix once as a standing convention for all example commands instead of repeating it on every line.

Collect version-sensitive notes (mcporter 0.13.10 fallback behavior, older/newer client timeout defaults) into a single 'Version-dependent behavior' or 'Legacy/older clients' section so the main flow stays timeless.

DimensionReasoningScore

Conciseness

The body is dense and almost entirely non-obvious operational detail, but it could be tightened: the long 'MCPORTER_CHROME_DEVTOOLS_RELAY_POLICY=require' prefix is repeated across roughly a dozen command lines, the relay-policy rationale is explained more than once, and time-sensitive material ('In mcporter 0.13.10', 'older clients', 'newer Chrome-specific outer defaults') sits inline rather than in a versions/deprecated section. This matches 'mostly efficient but includes some unnecessary explanation or could be tightened' (3), short of the 'minor instances' anchor (4).

3 / 5

Actionability

Fully executable, copy-paste-ready mcporter commands with real arguments and output flags, a concrete JSON config block to merge into the canonical definition, and specific failure signatures ('network-error' at the wrong endpoint, 'browser_owner_conflict', stale uid reports, blocking 'Allow remote debugging?' prompts) each paired with a remedy. This matches 'fully executable; copy-paste ready code or commands; specific examples cover the common cases'.

5 / 5

Workflow Clarity

The route is an explicit ordered fallback chain, setup/repair is a concrete checklist, and the 'Fail-Closed Readiness Proof' gives a five-condition numbered validation gate with commands. Error-recovery feedback loops are strong and explicit: a relay-policy error means report/repair instead of retrying, a stale uid means re-snapshot rather than retry, and an empty page list triggers a defined diagnosis sequence instead of escalation. This matches the top anchor including checklists and feedback loops.

5 / 5

Progressive Disclosure

The single SKILL.md is well-sectioned with clear headers, but there are no bundle files at all (references/, scripts/, assets/ are absent) and the ~290-line body inlines material that clearly belongs in separate files — daemon and timeout configuration, extension setup/repair procedures, and the legacy fallback. This matches 'some structure but could be better organized; content that should be separate is inline' (3), above the unstructured anchor (2) but below the 'appropriately split' anchors (4-5).

3 / 5

Total

16

/

20

Passed

Description

53%Weight 40%Scale 1-5

Based on the skill's description, can an agent find and select it at the right time? Clear, specific descriptions lead to better discovery.

The description is distinctive and technically precise about routing, but it describes tool selection rather than user-facing capabilities, omits any 'use when' trigger guidance, and lacks the natural terms (login, browser automation, screenshots) a user would say when needing this skill. Adding an explicit trigger clause and naming the concrete browser actions would lift specificity, completeness, and trigger quality together.

Suggestions

Add an explicit 'Use when...' clause, e.g. 'Use when the user asks to control, verify, or log into sites in their real signed-in Chrome, or when isolated test browsers would miss their cookies, SSO, or extensions.'

Name the concrete browser operations the skill actually performs (list/select tabs, navigate, click, fill, snapshot, evaluate JavaScript, screenshot) so the description states capabilities, not just internal tool routing.

Include natural trigger synonyms such as 'browser automation', 'log in / sign in', 'SSO', and 'live UI verification' that users would plausibly say verbatim.

DimensionReasoningScore

Specificity

The description names the domain ('Control the user's existing signed-in Chrome') and a concrete tool-routing order ('callable Codex Chrome plugin first, then OpenClaw extension-backed mcporter, with direct DevTools attachment only as a last fallback'), but it never lists the actual browser operations the skill performs (navigate, click, fill, snapshot, evaluate). It matches the anchor 'names domain and 1-2 concrete actions, but not comprehensive' rather than score 4, which requires several specific listed actions.

3 / 5

Completeness

The 'what' is clear (control the user's existing signed-in Chrome with a defined routing order), but there is no 'Use when...' clause or equivalent trigger guidance anywhere in the description. Per the rubric guideline, a missing 'Use when' clause caps completeness at 3.

3 / 5

Trigger Term Quality

Natural keywords like 'Chrome', 'signed-in', and 'DevTools' are present, but common phrasings a user would actually say are missing: 'browser automation', 'log in' / 'login', 'SSO', 'screenshot', 'web page'. This fits 'some relevant keywords but missing common variations or synonyms' (3), below the 'good keyword coverage' anchor (4).

3 / 5

Distinctiveness Conflict Risk

It carves a clear niche (the user's real, signed-in Chrome profile via named tooling) that is distinguishable from isolated-browser automation skills. It is not a 5 because the trigger boundary is not stated, leaving minor overlap risk with a generic browser/web-automation skill; it is well above anchor 3's 'could still overlap with similar skills'.

4 / 5

Total

13

/

20

Passed

Validation

100%

Checks the skill against the spec for correct structure and formatting. All validation checks must pass before discovery and implementation can be scored.

Validation — 16 / 16 Passed

Validation for skill structure

No warnings or errors.

Repository
steipete/agent-scripts
Reviewed

Table of Contents

Is this your skill?

If you maintain this skill, you can claim it as your own. Once claimed, you can manage eval scenarios, bundle related skills, attach documentation or rules, and ensure cross-agent compatibility.