Content
92%Weight 40%Scale 1-5Reviews the quality of instructions and guidance provided to agents. Good implementation is clear, handles edge cases, and produces reliable results.
A well-crafted analysis skill: the missing-case versus weak-assertion decision rule, per-mutator separating inputs, and the equivalent-mutant/undecidability treatment encode genuine domain expertise rather than generic instruction. Structure, progressive disclosure, and workflow sequencing are all strong; the only weakness is length, with a few sections that could be tightened without losing content.
Suggestions
Move the 'If a boundary test already exists and the mutant still survived' discussion out of the Step 4 example template block into its own short subsection, so the template stays a clean copy-paste shape.
Tighten the Step 5 prose: the Stryker/PIT citations and the score-convention guidance could be compressed into a compact list, cutting ~15 lines without losing the load-bearing rules.
The 'Limitations' section partially duplicates content already stated inline (Steps 2, 4, 5); consider merging the coverage-link and operator-coverage caveats into the reference file where the per-tool detail already lives.
| Dimension | Reasoning | Score |
|---|---|---|
Conciseness | The body is information-dense (per-mutator separating inputs, tool quirks, score conventions) with little padding, but at ~300 lines it has trimmable sections: the Step 4 example block embeds a discussion paragraph ('If a boundary test already exists...') inside the template, and the Step 5 prose could be tightened. Between anchors 4 and 5 — efficient with minor instances of over-explanation. | 4 / 5 |
Actionability | Fully executable guidance: copy-paste-ready TypeScript for the SurvivedMutant record, four worked mutation families each with concrete separating inputs and proposed fixes, and a complete worked example test (`expect(() => cart.addItem({ qty: 100 })).not.toThrow()`) plus an exact output-format template. This is an instruction-only skill and the guidance is fully actionable. | 5 / 5 |
Workflow Clarity | Clear five-step sequence (normalize → classify → heuristics → propose → equivalence) with an explicit rationale for step ordering, a four-item completeness checklist in Step 4 acting as validation, and error-recovery branching ('If a boundary test already exists and the mutant still survived, either the assertion does not observe the throw... or the production boundary is off by one'). The skill is explicitly read-only ('Never auto-rewrite tests'), so the destructive-operation cap does not apply. | 5 / 5 |
Progressive Disclosure | One reference file (references/tool-normalization.md), clearly signaled with two inline links and a summary of what it carries, exactly one level deep and verified to exist with substantive per-tool tables. The body keeps the workflow and offloads per-tool field mappings and operator-name tables appropriately. | 5 / 5 |
Total | 19 / 20 Passed |