CtrlK
BlogDocsLog inGet started
Tessl Logo

drive-desktop-cdp

Drive a running OpenWork desktop window over CDP from the shell. Evaluate JS, take screenshots, open a new session, send a prompt with a chosen model, wait for the run to finish. Use when checking a UI change by hand in a world or pnpm dev, or when reproducing a chat state.

75

Quality

94%

Does it follow best practices?

Run evals on this skill

Adds up to 20 points to the overall score

View guide
SecuritybySnyk

Passed

No findings from the security scan

SKILL.md
Quality
Evals
Security

Quality

Content

93%Weight 40%Scale 1-5

Reviews the quality of instructions and guidance provided to agents. Good implementation is clear, handles edge cases, and produces reliable results.

An exemplary lean operational skill: executable commands verified against the real bundle script, an explicit boundary note steering verdicts to a sibling skill, and safety rules covering real-state worlds. The only gap is the absence of error-recovery guidance for failure cases like a wait-idle timeout or a missing window.

DimensionReasoningScore

Conciseness

The ~36-line body contains zero concept explanations or padding — every line is a command, a path, or an operational rule ("Reload after state-heavy changes", "Never type into an existing session"). It assumes Claude's competence and every token earns its place, matching the 'lean and efficient' anchor exactly.

5 / 5

Actionability

The Commands block is fully executable, copy-paste-ready bash (`node $S new-session`, `node $S wait-idle 180`, `node $S shot /tmp/after.png`, `node $S eval '...'`) verified against the real scripts/cdp.mjs subcommands, and the sqlite3 fallback includes a complete query string. It covers the common cases for this skill end-to-end.

5 / 5

Workflow Clarity

The sequence (find window → export CDP_URL → new-session → send → wait-idle → shot/eval) is clear with inline ordering comments ("always start a fresh session first", "until the Run task button is back") and wait-idle acts as a synchronization checkpoint before screenshotting. It falls short of a 5 because there is no error-recovery guidance — nothing on what to do if wait-idle times out, no OpenWork page is found, or the send fails.

4 / 5

Progressive Disclosure

The skill is under 50 lines with no need for external reference files, and its sections (Find the window / Commands / Rules) are well-organized and easy to navigate. The single bundle file, scripts/cdp.mjs, is referenced by its exact working path, matching the actual bundle structure.

5 / 5

Total

19

/

20

Passed

Description

92%Weight 40%Scale 1-5

Based on the skill's description, can an agent find and select it at the right time? Clear, specific descriptions lead to better discovery.

A strong description: concrete, comprehensive capability list in third person with an explicit 'Use when' clause carrying natural triggers. The only improvement space is expanding CP jargon (e.g., Chrome DevTools Protocol) and a few more user-voice synonyms for the trigger terms.

DimensionReasoningScore

Specificity

The description lists multiple specific concrete actions — "Evaluate JS, take screenshots, open a new session, send a prompt with a chosen model, wait for the run to finish" — comprehensively covering the script's actual subcommands, in third-person voice. It matches the top anchor and is not merely 'several actions with minor gaps'.

5 / 5

Completeness

It explicitly answers both what ("Drive a running OpenWork desktop window over CDP from the shell" plus the five concrete actions) and when ("Use when checking a UI change by hand in a world or pnpm dev, or when reproducing a chat state") with concrete trigger phrases, exactly matching the top anchor.

5 / 5

Trigger Term Quality

Natural trigger phrases are present ("checking a UI change by hand", "reproducing a chat state", "screenshots", "send a prompt"), giving good keyword coverage. It falls below the comprehensive anchor because "CDP" is unexpanded jargon a user might not say, and common phrasings like "screenshot the desktop app" or "debug the UI" variants are missing.

4 / 5

Distinctiveness Conflict Risk

The niche is a specific app's desktop window driven over CDP — "OpenWork desktop window over CDP" — with triggers (manual UI check, chat-state reproduction) distinct from automated eval/spec testing referenced in the body. Minimal conflict risk with other skills.

5 / 5

Total

19

/

20

Passed

Validation

100%

Checks the skill against the spec for correct structure and formatting. All validation checks must pass before discovery and implementation can be scored.

Validation — 16 / 16 Passed

Validation for skill structure

No warnings or errors.

Repository
different-ai/openwork
Reviewed

Table of Contents

Is this your skill?

If you maintain this skill, you can claim it as your own. Once claimed, you can manage eval scenarios, bundle related skills, attach documentation or rules, and ensure cross-agent compatibility.