CtrlK
BlogDocsLog inGet started
Tessl Logo

computer-use

Drive native desktop apps through DeepChat's built-in Computer Use tools. Use when the user asks to operate, inspect, automate, or perform a GUI task in a real desktop application.

69

Quality

85%

Does it follow best practices?

Run evals on this skill

Adds up to 20 points to the overall score

View guide

SecuritybySnyk

Passed

No findings from the security scan

SKILL.md
Quality
Evals
Security

Quality

Content

88%Weight 40%Scale 1-5

Reviews the quality of instructions and guidance provided to agents. Good implementation is clear, handles edge cases, and produces reliable results.

A dense, highly operational skill body: the required tool loop, verification contract, error-recovery rules, and platform caveats are all stated with exact tool names and parameters rather than vague direction. Its only soft spots are a small amount of inline detail that could live in reference files and the inability to verify the four linked reference files against the bundle.

DimensionReasoningScore

Conciseness

The body is dense and largely non-redundant — it assumes competence, explains no general concepts, and every section carries operational content (e.g., "Omit cursor_theme during normal session setup", "Never send element_token: \"\"\"). Minor trimmable spots remain: the Agent Cursor theme-profile detail ("retired v1 themes with modifier artwork are not compatible... session badge rather than by theme modifier assets") and some verification guidance repeated across 'Required Loop' and 'Action Results and Verification' keep it below the lean-every-token-earns-its-place anchor of 5.

4 / 5

Actionability

Guidance is concrete and directly executable in the plugin's tool vocabulary: exact calls with parameters such as `start_session({ session, capture_scope: "auto" })`, `click({ pid, window_id, x, y, session })`, `zoom({ pid, window_id, x1, y1, x2, y2, session })`, plus enumerated refusal codes (`snapshot_id_required`, `stale_element_token`) with a precise retry-once recovery recipe. Specific examples cover the common cases, matching the fully actionable anchor.

5 / 5

Workflow Clarity

The 'Required Loop' is a clearly sequenced 9-step workflow with explicit validation checkpoints — snapshot before every action, verify after every action, a defined recovery loop ("take one fresh get_window_state and retry once with a token or index-plus-snapshot pair entirely from the new result"), and mandatory `end_session` cleanup including error paths. Feedback loops and postcondition checks are explicit throughout, matching the top anchor.

5 / 5

Progressive Disclosure

Structure is good: a 'Linked References' section cleanly signals four one-level-deep files with one-line descriptions (`README.md`, `WEB_APPS.md`, `RECORDING.md`, `TESTS.md`), and browser/recording detail is deferred to them. It falls short of 5 because none of the referenced files are present in the bundle to verify, and detailed contract material (the full Agent Cursor section, the complete ActionResult effect taxonomy) is inlined in SKILL.md rather than split out.

4 / 5

Total

18

/

20

Passed

Description

82%Weight 40%Scale 1-5

Based on the skill's description, can an agent find and select it at the right time? Clear, specific descriptions lead to better discovery.

A concise, third-person description with an explicit and well-phrased 'Use when' trigger clause covering both what and when. Its only weakness is that the capability statement stays at the level of one broad action rather than enumerating the concrete GUI operations the skill actually performs.

Suggestions

Enumerate 2-3 concrete capabilities in the description, e.g., 'Click, type, and inspect native desktop app windows' alongside the current 'Drive native desktop apps' framing.

Add one or two natural user phrasings as trigger synonyms (e.g., 'automate a desktop app', 'use a GUI application') to widen keyword coverage.

DimensionReasoningScore

Specificity

"Drive native desktop apps through DeepChat's built-in Computer Use tools" names the domain and a single concrete capability but does not enumerate the concrete actions the skill performs (clicking, typing, screenshots, window inspection), matching the 'names domain and 1-2 concrete actions, but not comprehensive' anchor. A 4 would require listing several specific actions like the PDF example's 'extracts text, fills forms, converts pages'.

3 / 5

Completeness

It explicitly answers both: what ("Drive native desktop apps through DeepChat's built-in Computer Use tools") and when ("Use when the user asks to operate, inspect, automate, or perform a GUI task in a real desktop application") with concrete trigger phrases, matching the top anchor. Not 4 because the 'when' clause is already explicit and specific, not merely present.

5 / 5

Trigger Term Quality

The 'when' clause supplies natural trigger words — "operate, inspect, automate, or perform a GUI task in a real desktop application" — that users would plausibly say. A few natural variants (e.g., 'click a button in an app', 'screenshot the app', named desktop apps) are missing, keeping it just below the comprehensive-with-synonyms anchor of 5.

4 / 5

Distinctiveness Conflict Risk

"Native desktop apps", "Computer Use tools", and "GUI task in a real desktop application" carve out a clear niche that is clearly distinguishable from browser/web-automation, file-processing, or generic scripting skills, with minimal overlap risk. It does not fall to 4 because there is no meaningful overlap with closely related skills implied by the wording.

5 / 5

Total

17

/

20

Passed

Validation

87%

Checks the skill against the spec for correct structure and formatting. All validation checks must pass before discovery and implementation can be scored.

Validation — 14 / 16 Passed

Validation for skill structure

CriteriaDescriptionResult

metadata_version

'metadata.version' is missing

Warning

frontmatter_unknown_keys

Unknown frontmatter key(s) found; consider removing or moving to metadata

Warning

Total

14

/

16

Passed

Repository
ThinkInAIXYZ/deepchat
Reviewed

Table of Contents

Is this your skill?

If you maintain this skill, you can claim it as your own. Once claimed, you can manage eval scenarios, bundle related skills, attach documentation or rules, and ensure cross-agent compatibility.