CtrlK
BlogDocsLog inGet started
Tessl Logo

blast-radius

Find what a change could break somewhere else before it ships, beyond the diff, and prove the one fact it's safe because of by running real code instead of writing it up. Use for 'blast radius of X', 'what could this break', or reviewing a small diff you don't trust.

68

Quality

82%

Does it follow best practices?

Run evals on this skill

Adds up to 20 points to the overall score

View guide

SecuritybySnyk

Passed

No findings from the security scan

SKILL.md
Quality
Evals
Security

Quality

Content

86%Weight 40%Scale 1-5

Reviews the quality of instructions and guidance provided to agents. Good implementation is clear, handles edge cases, and produces reliable results.

A tight, high-signal instruction skill: a lean body with no over-explanation, a mostly concrete workflow anchored by an evidence ladder and an explicit prove-it-or-mark-it-unproven validation step, and clean single-file organization appropriate to its size. The only soft spot is a handful of abstract hints in step 3 and the lack of a fix-and-retry loop that keeps actionability and workflow clarity just below top marks.

DimensionReasoningScore

Conciseness

The ~45-line body is lean and assumes Claude's competence throughout: it never explains what a diff, grep, or teardown is, and every line carries instruction weight (e.g. 'Listing the callers is not the job. The agent can grep those in a second. The job is the breakage grep won't show you.'). This matches anchor 5 ('Lean and efficient; assumes Claude's competence; every token earns its place'); not 4 because there is no padded or over-explanatory passage that would need trimming — the brief framing sentences ('A blast-radius writeup that sounds right is worthless') justify a rule rather than padding.

5 / 5

Actionability

For an instruction-only skill the guidance is largely concrete: a 5-level evidence ladder ('You ran it. A script or test that calls the real code and fails loud if you're wrong'), a recipe for step 4 ('one small script that imports the same library the app ships and calls the exact function you're worried about'), and a specific output checklist. Not 5 because a few directions stay abstract (e.g. 'Work out when things run: microtasks, unmount and teardown, Solid versus React' — hints rather than executable steps, and oddly specific to Solid/React); not 3 because nothing is pseudocode-level and the key operations (read diff, find safety fact, write and run a proof script, paste output) are directly executable.

4 / 5

Workflow Clarity

The 6 'Steps' are clearly sequenced (read change → find the safety fact → look where grep stops → assess risks → prove the fact → fan out for big changes) with an explicit validation checkpoint ('Write a script or test that runs the real code, run it, and paste what happened') and a failure path ('If you couldn't prove it, write unproven'). This fits anchor 4 ('Clear sequence with most checkpoints present; minor validation gaps'). Not 5 because there is no fix-and-retry feedback loop and the evidence ladder, the steps, and the handback checklist are three parallel structures the reader must merge themselves; not 3 because validation is explicitly present, not missing or implicit.

4 / 5

Progressive Disclosure

No bundle files exist (no references/, scripts/, or assets/), and the self-contained body is under 50 lines with well-organized sections ('Don't trust your own writeup', 'How sure are you', 'Steps', 'What to hand back') — squarely the simple-skill case the rubric says can score 5 on organization alone. All cross-references are to sibling skills (`how`, `why`, `unslop`, `arena`), not nested files, so there is no reference depth or navigation problem. Not 4 because nothing that belongs in a separate file is inlined and no organization gaps exist.

5 / 5

Total

18

/

20

Passed

Description

78%Weight 40%Scale 1-5

Based on the skill's description, can an agent find and select it at the right time? Clear, specific descriptions lead to better discovery.

A strong description that explicitly answers both what the skill does and when to use it, with natural, quotable trigger phrases. Its main weaknesses are modest breadth of action coverage and missing common synonyms (impact analysis, downstream effects) that would push trigger coverage to comprehensive.

Suggestions

Broaden the action coverage in the 'what' clause (e.g. 'assess each risk's likelihood and cost' or 'hand back a merge-ready checklist of risks and cleared items') to raise specificity from 1-2 concrete actions to several.

Add natural synonyms to the trigger list — 'impact analysis', 'what does this affect', 'downstream effects', 'ripple effects' — so users who don't say 'blast radius' still land on this skill.

Tighten 'prove the one fact it's safe because of by running real code instead of writing it up' — the embedded contrast makes the clause long; a crisper verb phrase would keep the what-clause concrete.

DimensionReasoningScore

Specificity

The description names the domain (finding downstream breakage from a change) and two concrete actions — "Find what a change could break somewhere else before it ships, beyond the diff" and "prove the one fact it's safe because of by running real code" — but coverage is not comprehensive (no mention of e.g. risk assessment or cleared-items output). This matches anchor 3 ("Names domain and 1-2 concrete actions, but not comprehensive"); not 4 because it does not list several distinct specific actions, and not 2 because the actions are concrete rather than minimal or generic.

3 / 5

Completeness

Both halves are explicit: the 'what' is "Find what a change could break somewhere else before it ships... and prove the one fact it's safe because of by running real code", and the 'when' is "Use for 'blast radius of X', 'what could this break', or reviewing a small diff you don't trust". This matches anchor 5 (clearly and explicitly answers both what AND when with concrete trigger phrases). Not 4 because the 'when' clause is fully explicit with quoted trigger phrases rather than merely present.

5 / 5

Trigger Term Quality

Trigger phrases "blast radius of X", "what could this break", and "reviewing a small diff you don't trust" are natural things a user would actually say, giving good keyword coverage. Not 5 because common synonyms and variations are missing (e.g. "impact analysis", "what does this affect", "downstream effects", "ripple effect"); not 3 because the included phrases go beyond 'some relevant keywords' — they read like real user utterances, including the exact trigger phrase 'blast radius of X'.

4 / 5

Distinctiveness Conflict Risk

The 'blast radius of X' / 'what could this break' framing carves out a clear niche with distinct triggers, and the emphasis on proving safety facts by running code differentiates it from plain review skills. Not 5 because "reviewing a small diff you don't trust" has minor overlap with general code-review/diff-review skills; not 3 because the primary triggers are specific and unlikely to fire for unrelated skills.

4 / 5

Total

16

/

20

Passed

Validation

100%

Checks the skill against the spec for correct structure and formatting. All validation checks must pass before discovery and implementation can be scored.

Validation — 16 / 16 Passed

Validation for skill structure

No warnings or errors.

Repository
cursor/plugins
Reviewed

Table of Contents

Is this your skill?

If you maintain this skill, you can claim it as your own. Once claimed, you can manage eval scenarios, bundle related skills, attach documentation or rules, and ensure cross-agent compatibility.