CtrlK
BlogDocsLog inGet started
Tessl Logo

stress-plan

Use when an implementation plan has just been written and before any code is written — hands the plan to a second model in a fresh context to attack it, then requires every finding to be closed or rejected in writing. Invoked by new-feat at the end of its planning step and by orchestrate once per phase, and worth running on its own for any plan whose assumptions have not been checked against the code.

68

Quality

82%

Does it follow best practices?

Run evals on this skill

Adds up to 20 points to the overall score

View guide

SecuritybySnyk

Passed

No findings from the security scan

The canonical home for this skill is stress-plan in englishstreetventures/osn

SKILL.md
Quality
Evals
Security

Quality

Content

81%Weight 40%Scale 1-5

Reviews the quality of instructions and guidance provided to agents. Good implementation is clear, handles edge cases, and produces reliable results.

The body is an exceptionally actionable, well-gated workflow with a copy-paste-ready dispatch prompt and explicit validation checkpoints. Its main cost is token weight: the anecdotal justifications, while valuable, are written at narrative length and could be tightened substantially.

Suggestions

Compress the five war stories (vite minify, sourcemap, GDPR batch, six-apps issue, double-written build dir) into one-line parenthetical evidence after the check they justify — e.g. 'cite the reader, or the value ships as a no-op (a plan once set `minify: true` and the framework hard-overrode it)'.

Move the 13-line attacker prompt template and the Step 1 checklist heuristics into a `references/` file (e.g. `attacker-prompt.md`), keeping SKILL.md as the overview plus dispatch protocol — this would also raise the progressive-disclosure score.

Trim Step 0's narrative framing ('When in doubt it runs. The stage costs one round trip; the failure it catches costs the branch') to a single directive sentence; the reasoning is already implied by the never-skip list.

DimensionReasoningScore

Conciseness

The prose assumes Claude's competence (no basic concepts explained) and there is no padding, but it carries five multi-line war stories (the vite `minify` no-op, the published `sourcemap`, the GDPR erasure batch, the six-apps issue premise, the two-process build measurement), each of which could be compressed to a clause while keeping its evidentiary value. This is 'mostly efficient but could be tightened' rather than 'efficient with minor trims', since the anecdotes together occupy a substantial fraction of the body.

3 / 5

Actionability

The attacker prompt is given verbatim and copy-paste ready ("Read `<plan path>`. Attack it... Write `<findings path>` and reply with the path only"), with concrete paths (`.claude/agents/attacker.md`, `PLAN-REVIEW.md`, `docs/superpowers/plans/`, `tools/pr-metrics/index.ts`, `resolveDispatchBranch`), an explicit `TASK-BRANCH:` dispatch protocol, and enumerated skip/never-skip criteria. As an instruction-only skill its guidance is fully executable and covers the common cases.

5 / 5

Workflow Clarity

Steps 0–4 are clearly sequenced with explicit validation gates ("No implementation starts while a finding is open. Silence is not a rejection."), a hard three-outcome enumeration for each finding (fix / reject in writing / escalate), evidence-based rejection checking ("Check the claim's numbers against the artefact"), and explicit re-run criteria. Error-recovery feedback loops are present throughout, matching the top anchor.

5 / 5

Progressive Disclosure

No bundle directories exist, and the body's structure is clear with well-labeled sections and concrete pointers to real repo artefacts (`.claude/agents/attacker.md`, `tools/pr-metrics/index.ts`). It is not a 5 because at ~140 lines with an inline 13-line attacker prompt and detailed check heuristics, some material could live in a reference file; the >50-line simple-skill exception does not apply.

4 / 5

Total

17

/

20

Passed

Description

83%Weight 40%Scale 1-5

Based on the skill's description, can an agent find and select it at the right time? Clear, specific descriptions lead to better discovery.

The description explicitly answers both what and when with concrete trigger language, and carves out a distinct niche. Trigger-term coverage is good but could add a few natural synonyms users might say.

DimensionReasoningScore

Specificity

Quotes: "hands the plan to a second model in a fresh context to attack it" and "requires every finding to be closed or rejected in writing" name the domain and several concrete, specific actions. Not a 5 because coverage is limited to the attack-and-close loop (e.g., what the findings report contains or what 'attack' checks for is left to the body), and not a 3 because the actions listed are specific and more than minimal.

4 / 5

Completeness

Both halves are explicit: "Use when an implementation plan has just been written and before any code is written" answers 'when' with concrete trigger phrasing, and the attack-then-close-findings mechanism answers 'what'. Not a 4 because the 'when' clause is already concrete and specific rather than merely present.

5 / 5

Trigger Term Quality

Natural phrases like "when an implementation plan has just been written", "before any code is written", and "any plan whose assumptions have not been checked" are terms a user would plausibly say. Missing a few natural synonyms ("plan review", "stress-test the plan", "critique the design"), which keeps it below the comprehensive anchor at 5 but comfortably above the sparse keyword anchor at 3.

4 / 5

Distinctiveness Conflict Risk

The niche is distinct — adversarial second-model review of a plan before implementation — and it names its calling skills ("Invoked by new-feat... and by orchestrate"), so it is unlikely to fire for unrelated skills. Minor overlap risk with generic plan-writing or code-review skills from phrases like "worth running on its own for any plan", keeping it just below the minimal-conflict anchor at 5.

4 / 5

Total

17

/

20

Passed

Validation

100%

Checks the skill against the spec for correct structure and formatting. All validation checks must pass before discovery and implementation can be scored.

Validation — 16 / 16 Passed

Validation for skill structure

No warnings or errors.

Repository
englishstreetventures/englishstreetventures.com
Reviewed

Table of Contents

Is this your skill?

If you maintain this skill, you can claim it as your own. Once claimed, you can manage eval scenarios, bundle related skills, attach documentation or rules, and ensure cross-agent compatibility.