CtrlK
BlogDocsLog inGet started
Tessl Logo

computer-use-agents

Build AI agents that interact with computers like humans do - viewing screens, moving cursors, clicking buttons, and typing text. Covers Anthropic's Computer Use, OpenAI's Operator/CUA, and open-source alternatives.

58

Quality

68%

Does it follow best practices?

Run evals on this skill

Adds up to 20 points to the overall score

View guide

SecuritybySnyk

Low

Low-risk findings worth noting

Fix and improve this skill with Tessl

tessl review fix ./skills/computer-use-agents/SKILL.md
SKILL.md
Quality
Evals
Security

Quality

Content

65%Weight 40%Scale 1-5

Reviews the quality of instructions and guidance provided to agents. Good implementation is clear, handles edge cases, and produces reliable results.

A well-structured thin-router SKILL.md: it is lean, points to a single real, well-organized reference one level deep, and includes concrete trigger and limitation sections. Its weaknesses are that the body itself carries no executable or domain-specific content (all operational detail is delegated) and validation checkpoints for this critical-risk skill are mandated by reference rather than surfaced, and the opening paragraph duplicates the frontmatter description.

Suggestions

Surface 1-2 of the guide's critical safety rules or validation commands directly in the body (e.g., the sandboxing prerequisite or a verify-after-action checkpoint) so the mandatory requirements are visible before the reference is loaded.

Add a short quick-start snippet (e.g., the minimal action-loop or a pointer to the guide's key section anchors) so the body is not purely a router with no executable content.

Trim the opening paragraph that restates the frontmatter description verbatim, keeping only the added sentence about the critical focus on sandboxing and security.

DimensionReasoningScore

Conciseness

The ~38-line body is efficient — a short overview, a clear pointer to the detailed guide, a trigger list, and limitations — with almost no over-explanation of things Claude already knows. It is a 4 rather than 5 because the opening paragraph restates the frontmatter description nearly verbatim ("Build AI agents that interact with computers like humans do - viewing screens, moving cursors, clicking buttons, and typing text"), which is a trimmable duplication, and not a 3 because there is no padded tutorial content anywhere else.

4 / 5

Actionability

The body gives directive guidance ("Read the detailed guide before executing this skill", "Treat its safety, prerequisites, and validation requirements as mandatory", "Stop and ask for clarification if required inputs, permissions, safety boundaries, or success criteria are missing") but contains no concrete code, commands, or domain-specific operational steps — all executable material is delegated to the reference. This matches the "some concrete guidance but incomplete / missing key details" anchor: above level 2 because the routing instructions and stop conditions are specific and unambiguous, below level 4 because nothing in the body itself is executable.

3 / 5

Workflow Clarity

A usage sequence is present (trigger match → read guide with focused or complete loading → execute under mandatory safety/validation), and validation is explicitly mandated ("Treat its safety, prerequisites, and validation requirements as mandatory") but no actual checkpoint is shown in the body — they are all delegated to the reference. For a skill marked risk: critical, this matches "checkpoints missing or implicit" and the rubric's cap of 3 for risky operations without concrete validation steps; it is above level 2 because the sequence that exists is coherent and the validation mandate is explicit.

3 / 5

Progressive Disclosure

The body is a clean overview with one clearly signaled, one-level-deep reference: the link to references/detailed-guide.md is real (a 2158-line, well-sectioned guide) and the body explains how to load it ("For focused work, load the relevant sections; for end-to-end work, read the guide completely"), and the guide itself contains no further nested file references. This matches the "clear overview with well-signaled one-level-deep references" anchor exactly, with nothing to distinguish it from a 4.

5 / 5

Total

15

/

20

Passed

Description

71%Weight 40%Scale 1-5

Based on the skill's description, can an agent find and select it at the right time? Clear, specific descriptions lead to better discovery.

A strong, concrete description that names specific actions and specific technologies in third-person imperative voice with no fluff. Its one structural weakness is the missing "Use when" trigger clause, which caps completeness and leaves natural synonyms (GUI automation, desktop automation, browser agent) out of the description itself.

Suggestions

Append an explicit trigger clause, e.g. "Use when the user mentions computer use, desktop automation, GUI automation, or browser agents, or asks for an agent that controls a screen, mouse, or keyboard."

Include a few more natural synonyms users would say ("GUI automation", "desktop automation", "RPA with AI", "browser agent") directly in the description rather than only in the body's When to Use section.

State the critical-risk safety posture briefly ("with mandatory sandboxing and security controls") so the description also signals the skill's guardrails.

DimensionReasoningScore

Specificity

The description lists four concrete actions ("viewing screens, moving cursors, clicking buttons, and typing text") plus explicit technology coverage ("Anthropic's Computer Use, OpenAI's Operator/CUA, and open-source alternatives"), matching the comprehensive multi-action anchor. It is not a 4 because coverage of the core interaction primitives plus the ecosystem landscape leaves only trivial gaps, and not lower because no part of it is vague or generic.

5 / 5

Completeness

The "what" is clear and concrete, but there is no "Use when..." clause or equivalent explicit trigger guidance in the description, capping completeness at 3 per the judging guidelines. It is not a 4 because the "when" is entirely absent rather than merely under-specified, and not a 2 because the "what" half is strong.

3 / 5

Trigger Term Quality

Natural phrases users would say are present ("computer use", "Operator", "AI agents", "screens", "cursors"), giving good keyword coverage. It is a 4 rather than 5 because common natural synonyms like "GUI automation", "desktop automation", "RPA", or "browser agent" are absent from the description (they only appear in the body's When to Use list), and rather than 3 because the included terms are exactly what a user with this need would say.

4 / 5

Distinctiveness Conflict Risk

Naming Anthropic's Computer Use and OpenAI's Operator/CUA carves out a mostly distinct niche with low confusion risk against unrelated skills. It is a 4 rather than 5 because the domain overlaps with closely related browser-automation, GUI-testing, and general agent-building skills that share triggers like "browser agent" or "automation agent".

4 / 5

Total

16

/

20

Passed

Validation

93%

Checks the skill against the spec for correct structure and formatting. All validation checks must pass before discovery and implementation can be scored.

Validation — 15 / 16 Passed

Validation for skill structure

CriteriaDescriptionResult

frontmatter_unknown_keys

Unknown frontmatter key(s) found; consider removing or moving to metadata

Warning

Total

15

/

16

Passed

Repository
sickn33/agentic-awesome-skills
Reviewed

Table of Contents

Is this your skill?

If you maintain this skill, you can claim it as your own. Once claimed, you can manage eval scenarios, bundle related skills, attach documentation or rules, and ensure cross-agent compatibility.