CtrlK
BlogDocsLog inGet started
Tessl Logo

agent-first-screenshots

Agent-first screenshots — an agent drives the real app via CDP and produces clean, defect-free product screenshots (newsletters, landing pages, social, decks, PR). Dual-channel verification (DOM + pixels + vision) in a capture loop. Use for any "take/redo screenshots of the app" task.

70

Quality

86%

Does it follow best practices?

Run evals on this skill

Adds up to 20 points to the overall score

View guide

SecuritybySnyk

Passed

No findings from the security scan

SKILL.md
Quality
Evals
Security

Quality

Content

88%Weight 40%Scale 1-5

Reviews the quality of instructions and guidance provided to agents. Good implementation is clear, handles edge cases, and produces reliable results.

A well-engineered, feedback-driven capture workflow: tight prose, calibrated numeric gates, a mandatory vision check, and real bundled scripts that exist and match their descriptions. The only notable gap is that script invocation details are left entirely to the script files rather than given a one-line usage pointer in the body.

DimensionReasoningScore

Conciseness

Lean and imperative throughout — every section carries novel domain judgment ('Only remove, never add', 'In DOM but not in pixels | Clipped/blank render — trust variance + vision') and numeric thresholds ('variance > ~200', 'bgRatio > 0.97'), with no explanations of concepts Claude already knows and no padding.

5 / 5

Actionability

Concrete and executable in most places: exact CDP specifics ('1440x900 at deviceScaleFactor: 2, wait ~1s and re-verify', 'location.reload()'), calibrated thresholds, a copy-paste JSON vision rubric, and two real bundled scripts. Not 5 because the body never shows how to invoke the scripts (e.g. a 'node capture-verify.mjs ...' usage line) or how to establish the CDP connection.

4 / 5

Workflow Clarity

The per-shot sequence is explicit ('location.reload(), wait for full render, navigate through the UI, minimal leaf cleanup, verify, capture') with three validation gates and genuine feedback loops ('If it fails: reload and redo — never fix with more CSS'; rejected vision verdicts route to 'Diagnose, fix the root cause... recapture'), reinforced by a symptom→fix failure table. The destructive/batch cap does not apply — this is neither.

5 / 5

Progressive Disclosure

The ~95-line body is well-sectioned and correctly splits the deterministic verification logic into two real one-level-deep scripts, each with a clear purpose line ('verify an existing PNG' vs 'capture at 2x + verify regions + save only on pass'). Not 5 because the script pointers are plain backtick mentions rather than links, and no usage example or arguments summary is surfaced in the body.

4 / 5

Total

18

/

20

Passed

Description

83%Weight 40%Scale 1-5

Based on the skill's description, can an agent find and select it at the right time? Clear, specific descriptions lead to better discovery.

A strong description: concrete, third-person, with an explicit 'Use for...' trigger containing natural user phrasing. Its main weaknesses are a thin set of trigger synonyms and a slight labeling inconsistency ('Dual-channel' over three channels).

Suggestions

Add natural trigger synonyms such as 'capture', 'snapshot', or 'marketing/product images' alongside 'take/redo screenshots of the app'.

Fix the 'Dual-channel verification (DOM + pixels + vision)' inconsistency (three channels are listed) — say 'Three-channel verification' or drop the count.

Consider a brief boundary clause in the description (e.g. 'not for e2e pass/fail evidence') to reduce overlap with testing skills.

DimensionReasoningScore

Specificity

Lists several concrete actions — 'drives the real app via CDP', 'produces clean, defect-free product screenshots (newsletters, landing pages, social, decks, PR)', 'Dual-channel verification (DOM + pixels + vision) in a capture loop' — with named output types. Not anchor 5 because coverage has minor gaps: the mechanism stays jargon-lean and 'Dual-channel' confusingly labels three listed channels.

4 / 5

Completeness

Explicitly answers both: what ('an agent drives the real app via CDP and produces clean, defect-free product screenshots... Dual-channel verification... in a capture loop') and when ('Use for any "take/redo screenshots of the app" task') with a concrete quoted trigger phrase, in third-person voice — a direct match for the anchor-5 example.

5 / 5

Trigger Term Quality

'screenshots' and the quoted 'take/redo screenshots of the app' are phrases a user would naturally say, giving good keyword coverage. Not 5 because common synonyms like 'capture', 'snapshot', 'marketing images', or 'product images' are absent; not 3 because the core natural phrases are present and well-chosen.

4 / 5

Distinctiveness Conflict Risk

'take/redo screenshots of the app' carves a clear niche unlikely to grab unrelated skills. Not 5 because the description alone doesn't delimit it from e2e-testing or screenshot-diff skills (that boundary lives only in the body's 'NOT e2e evidence — use the fraimz skill' note), leaving minor overlap risk.

4 / 5

Total

17

/

20

Passed

Validation

100%

Checks the skill against the spec for correct structure and formatting. All validation checks must pass before discovery and implementation can be scored.

Validation — 16 / 16 Passed

Validation for skill structure

No warnings or errors.

Repository
Devin-AXIS/iPolloWork
Reviewed

Table of Contents

Is this your skill?

If you maintain this skill, you can claim it as your own. Once claimed, you can manage eval scenarios, bundle related skills, attach documentation or rules, and ensure cross-agent compatibility.