CtrlK
BlogDocsLog inGet started
Tessl Logo

testland/chromatic-visual-regression-testing

Authors and runs Chromatic visual tests on Storybook, Playwright, or Cypress projects via the `chromatic` CLI; configures baselines, TurboSnap, UI Review, and CI gating; reads exit codes for change-vs-error classification. Use when the project ships visual regression coverage to Chromatic Cloud.

65

Quality

82%

Does it follow best practices?

Run evals on this skill

Adds up to 20 points to the overall score

View guide
SecuritybySnyk

Passed

No findings from the security scan

Overview
Quality
Evals
Security
Files

Quality

Content

75%Weight 40%Scale 1-5

Reviews the quality of instructions and guidance provided to agents. Good implementation is clear, handles edge cases, and produces reliable results.

A dense, highly actionable single-file skill: executable commands, complete CI yaml, and semantically rich exit-code documentation. Its main weaknesses are the missing story-authoring example, no error-recovery guidance tied to exit codes, and all reference-grade detail living inline rather than in bundle files.

Suggestions

Add a minimal Storybook story example showing a parameters.chromatic block (e.g. viewports and delay) so the primary authoring mode is executable too.

Move the full flag table and exit-code catalog to a references/cli.md file, keeping SKILL.md as an overview with well-signaled one-level-deep pointers.

Add brief error-recovery guidance per exit-code family (e.g. on 21-23 Storybook build failures, fix the build locally with --dry-run before re-running).

DimensionReasoningScore

Conciseness

The body is dense with operational detail (flag table, exit-code table, config JSON, CI yaml) and assumes Claude already knows Storybook/Playwright. Minor over-explanation remains, such as the 'built by the team behind Storybook' context line and TurboSnap being explained in three separate places.

4 / 5

Actionability

Install, first-run, flags, config file, and a complete GitHub Actions workflow are copy-paste ready, but the primary authoring path ('Authoring stories (Storybook mode)') describes parameters.chromatic knobs (viewports, delay, pauseAnimationAtEnd) without any story code example. Not 5: a common case lacks executable code.

4 / 5

Workflow Clarity

Install → baseline-establishing first run → author → run → exit-code classification → CI gating is clearly sequenced, and the exit-code table plus the 'fail on exit != 0 (or != 0,1)' guidance act as checkpoints. No destructive/batch cap applies, but there is no explicit error-recovery loop (e.g. what to do on exit 21-23), keeping it below 5.

4 / 5

Progressive Disclosure

Sections are well organized and external doc references are clearly signaled, but the skill is a single 227-line file: the full CLI flag table, exit-code list, and TurboSnap detail are all inlined in SKILL.md rather than split into one-level-deep reference files. Not 5: no reference-file structure; not 3: navigation and signaling are good.

4 / 5

Total

16

/

20

Passed

Description

80%Weight 40%Scale 1-5

Based on the skill's description, can an agent find and select it at the right time? Clear, specific descriptions lead to better discovery.

A strong, specific, third-person description with concrete actions and an explicit use-when clause. Its main gaps are missing natural synonyms for visual testing terms and a when-clause that triggers on project state rather than user intent.

Suggestions

Broaden the 'Use when' clause with user-mention triggers, e.g. 'Use when the project ships visual regression coverage to Chromatic Cloud, or when the user mentions Chromatic, visual regression, or snapshot review.'

Add one or two natural synonyms such as 'visual snapshots' or 'UI review' phrasing users are likely to say when asking for this skill.

DimensionReasoningScore

Specificity

"Authors and runs Chromatic visual tests... configures baselines, TurboSnap, UI Review, and CI gating; reads exit codes for change-vs-error classification" lists multiple specific concrete actions in third person with comprehensive coverage, matching the top anchor rather than the 'minor gaps' level below.

5 / 5

Completeness

Both a clear 'what' (author/run/configure/read exit codes) and an explicit 'Use when the project ships visual regression coverage to Chromatic Cloud' clause are present. Not 5: the when-clause covers only project state and lacks concrete user-mention trigger phrases (e.g. 'when the user mentions visual regression or snapshots').

4 / 5

Trigger Term Quality

Strong natural keywords (Chromatic, visual tests, visual regression, Storybook/Playwright/Cypress, TurboSnap, CI gating, Chromatic Cloud), but common user phrasings like "visual snapshots", "screenshots", or "UI tests" are absent. Not 3: coverage is genuinely good; not 5: a few natural synonyms are missing.

4 / 5

Distinctiveness Conflict Risk

The Chromatic-specific niche is clear, but naming 'Storybook, Playwright, or Cypress projects' creates minor overlap risk with generic testing skills for those frameworks. Not 5: the framework names slightly widen the trigger surface.

4 / 5

Total

17

/

20

Passed

Validation

100%

Checks the skill against the spec for correct structure and formatting. All validation checks must pass before discovery and implementation can be scored.

Validation — 16 / 16 Passed

Validation for skill structure

No warnings or errors.

Reviewed

Table of Contents