CtrlK
BlogDocsLog inGet started
Tessl Logo

agent-browser

Browser automation CLI for AI agents. Use when the user needs to interact with websites, including navigating pages, filling forms, clicking buttons, taking screenshots, extracting data, testing web apps, or automating any browser task. Triggers include requests to "open a website", "fill out a form", "click a button", "take a screenshot", "scrape data from a page", "test this web app", "login to a site", "automate browser actions", or any task requiring programmatic web interaction.

68

Quality

85%

Does it follow best practices?

Run evals on this skill

Adds up to 20 points to the overall score

View guide
SecuritybySnyk

Critical

Do not install without reviewing

SKILL.md
Quality
Evals
Security

Quality

Content

71%Weight 40%Scale 1-5

Reviews the quality of instructions and guidance provided to agents. Good implementation is clear, handles edge cases, and produces reliable results.

The body is highly actionable, with executable commands and clear workflow patterns throughout, and it links to a real, well-organized set of reference files. Its weaknesses are redundancy (inline command reference and repeated session sections that duplicate both each other and the bundle references) and broken template links that undermine navigation.

Suggestions

Fix the progressive-disclosure defect: either add the three templates/*.sh files to the bundle or remove the "Ready-to-Use Templates" table and invocation examples, since those paths currently 404.

Trim the "Essential Commands" section to a short core set and point to references/commands.md for the full reference, removing the duplication.

Merge the "Parallel Sessions" and "Session Persistence" patterns into the "Session Management and Cleanup" section so each topic appears once.

DimensionReasoningScore

Conciseness

Most content is agent-specific command knowledge Claude cannot know, but there is notable redundancy: the "Essential Commands" block (~55 lines) largely duplicates references/commands.md, and "Parallel Sessions" + "Session Persistence" restate the later "Session Management and Cleanup" section, while "Timeouts and Slow Pages" repeats wait flags already listed above. This fits "mostly efficient but ... could be tightened" rather than the noticeably padded anchor 2, but not the "efficient; minor instances" anchor 4 given the volume of duplication.

3 / 5

Actionability

Nearly every section is copy-paste-ready executable commands — e.g. the worked login flow `agent-browser open https://example.com/form` → `snapshot -i` → `fill @e1` → `click @e3` → `wait --load networkidle` — and concrete recipes cover the common cases (form submission, auth vault, extraction, diffing, iOS). It fully matches "fully executable; copy-paste ready code or commands; specific examples cover the common cases".

5 / 5

Workflow Clarity

The "Core Workflow" gives a clear numbered sequence (Navigate → Snapshot → Interact → Re-snapshot) with an explicit checkpoint rule ("Refs ... are invalidated when the page changes. Always re-snapshot after"), and the Diffing section supplies a verification step ("Use `diff snapshot` after performing an action to verify it had the intended effect"). It falls short of anchor 5 because error-recovery loops (e.g. what to do when a ref is stale or a wait times out mid-flow) are deferred to reference files rather than stated as feedback loops, leaving "minor validation gaps" at anchor 4.

4 / 5

Progressive Disclosure

The "Deep-Dive Documentation" table with a "When to Use" column is a well-signaled one-level-deep structure and all 7 referenced .md files exist, but the body also links three scripts under templates/ (e.g. [templates/form-automation.sh](templates/form-automation.sh)) that do not exist in the bundle, and the inlined full command reference is content that belongs in the already-present commands.md. Per the guideline to score against the actual bundle structure, 3 of 10 referenced paths are broken, fitting anchor 3 ("content that should be separate is inline") better than anchor 4's "minor organization gaps".

3 / 5

Total

15

/

20

Passed

Description

96%Weight 40%Scale 1-5

Based on the skill's description, can an agent find and select it at the right time? Clear, specific descriptions lead to better discovery.

A strong description that clearly states what the tool does and when to use it, with concrete capability enumeration and an unusually good set of natural-language trigger phrases. The only weakness is that a few trigger terms (screenshot, click a button) are not browser-qualified and carry minor conflict risk with adjacent skills.

Suggestions

Qualify ambiguous trigger phrases, e.g. "take a screenshot (of a web page)" or "click a button (on a website)", to reduce overlap with generic screenshot/UI-automation skills.

Trim the redundant catch-all "or any task requiring programmatic web interaction", which restates the preceding trigger list without adding discrimination.

DimensionReasoningScore

Specificity

The description enumerates concrete, comprehensive actions — "navigating pages, filling forms, clicking buttons, taking screenshots, extracting data, testing web apps, or automating any browser task" — with no vague filler, matching the anchor for multiple specific concrete actions with comprehensive coverage. Nothing is missing for the browser-automation domain, so the anchor below ("several specific actions; minor gaps") does not fit.

5 / 5

Completeness

Both questions are answered explicitly: "what" is "Browser automation CLI for AI agents" with enumerated capabilities, and "when" is a double explicit trigger clause — "Use when the user needs to interact with websites" plus "Triggers include requests to ..." with concrete phrases. This matches the top anchor exactly; the score-4 anchor ("'when' could be more explicit") is clearly surpassed.

5 / 5

Trigger Term Quality

It quotes natural user phrasings directly — "open a website", "fill out a form", "click a button", "take a screenshot", "scrape data from a page", "test this web app", "login to a site", "automate browser actions" — covering synonyms and common variations users would actually say. This exceeds the "good keyword coverage; a few natural terms missing" anchor.

5 / 5

Distinctiveness Conflict Risk

The niche is clear and browser-scoped ("Browser automation CLI", "interact with websites"), but several bare trigger phrases — "take a screenshot", "click a button" — could plausibly fire for non-browser screenshot or UI-automation skills, which is minor overlap with closely related skills. It is well above anchor 3 ("could still overlap with similar skills") because the overall framing consistently scopes to web/browser interaction, but not a clean fit for anchor 5's "minimal conflict risk".

4 / 5

Total

19

/

20

Passed

Validation

81%

Checks the skill against the spec for correct structure and formatting. All validation checks must pass before discovery and implementation can be scored.

Validation — 13 / 16 Passed

Validation for skill structure

CriteriaDescriptionResult

skill_md_line_count

SKILL.md is long (542 lines); consider splitting into references/ and linking

Warning

allowed_tools_field

'allowed-tools' contains unusual tool name(s)

Warning

relative_links

Relative link issues: 3 missing

Warning

Total

13

/

16

Passed

Repository
ZHangZHengEric/Sage
Reviewed

Table of Contents

Is this your skill?

If you maintain this skill, you can claim it as your own. Once claimed, you can manage eval scenarios, bundle related skills, attach documentation or rules, and ensure cross-agent compatibility.