CtrlK
BlogDocsLog inGet started
Tessl Logo

subagent-driven-development

Use when executing implementation plans with independent tasks in the current session

48

Quality

53%

Does it follow best practices?

Run evals on this skill

Adds up to 20 points to the overall score

View guide

SecuritybySnyk

Passed

No findings from the security scan

Fix and improve this skill with Tessl

tessl review fix ./plugins/superpowers/skills/subagent-driven-development/SKILL.md

The canonical home for this skill is subagent-driven-development in obra/superpowers

SKILL.md
Quality
Evals
Security

Quality

Content

70%Weight 40%Scale 1-5

Reviews the quality of instructions and guidance provided to agents. Good implementation is clear, handles edge cases, and produces reliable results.

A mature, highly actionable process document with an exemplary review/fix feedback loop and validation checkpoints, but it is monolithic: an already-long body carries content that belongs in references, and nearly half its referenced files are missing from the bundle. Tightening the repetition and shipping or inlining the templates would lift it further.

Suggestions

Ship the three missing template files (implementer-prompt.md, task-reviewer-prompt.md, re-review-prompt.md) or inline their essential contract text — three of six referenced paths resolve to nothing in the bundle.

Move self-contained blocks (Model Selection detail, waiting-on-subagents guidance, the Example Workflow) into one-level-deep reference files to slim the body and reduce repeated statements of the rulings rule.

Replace the dot-graph edge list (which restates every node label) with a compact numbered loop description, keeping the node text once.

DimensionReasoningScore

Conciseness

Mostly dense operational guidance Claude does not already know (ledger formats, fix-round caps, BASE-not-HEAD~1 rules), but at ~570 lines it could be tightened: the rulings rule is restated in "Rulings, not stalls", Setup, the fix loop, and Finish; and the dot-graph edge list repeats every node label verbatim. This fits 'mostly efficient but includes some unnecessary explanation or could be tightened' — not 4 because the repetition and the duplicated graph text are more than minor trims.

3 / 5

Actionability

Concrete and executable throughout: exact commands ("scripts/task-brief PLAN_FILE N", "git rev-parse HEAD", "scripts/review-package PLAN_FILE BASE HEAD"), verbatim ledger line formats, a four-status report taxonomy, and a worked example. It is not 5 because three of the referenced dispatch templates (implementer-prompt.md, task-reviewer-prompt.md, re-review-prompt.md) are absent from the bundle, leaving the actual prompt text unstated — a minor but real gap.

4 / 5

Workflow Clarity

The process is fully sequenced (Setup → per-task dispatch → report handling → review → fix loop → completion → final review → finish) with explicit validation gates and a textbook feedback loop: review findings → fix round → scoped re-review → breaker adjudication at the 5-round cap. Destructive actions (rm -rf workspace) are guarded by explicit preconditions, and the Common Rationalizations table plus mandatory ledger lines act as checklists. This matches the top anchor.

5 / 5

Progressive Disclosure

The body has good section structure and correctly signals the three bundle scripts with their usage, but three of the six referenced paths (implementer-prompt.md, task-reviewer-prompt.md, re-review-prompt.md) do not exist in the bundle, and heavy content (Model Selection detail, waiting-on-subagents guidance, the 60-line example workflow) is inlined that belongs in reference files. This fits 'some structure but could be better organized... content that should be separate is inline' better than the 4 anchor.

3 / 5

Total

15

/

20

Passed

Description

36%Weight 40%Scale 1-5

Based on the skill's description, can an agent find and select it at the right time? Clear, specific descriptions lead to better discovery.

The description is a pure trigger clause: explicit and reasonably natural about when to fire, but completely silent on what the skill does. Users (and Claude) selecting among similar plan-execution skills get no capability signal to disambiguate.

Suggestions

State the 'what' before the 'when': e.g., "Execute an implementation plan by dispatching a fresh implementer subagent per task, reviewing each task's spec compliance and code quality, and running a final whole-branch review.

Add natural trigger variations and synonyms such as "run the plan", "work through the plan's tasks", or "delegate the plan's tasks to subagents" so more user phrasings match.

Name the distinguishing mechanism (subagent-per-task with review gates) to reduce overlap with generic plan-execution skills.

DimensionReasoningScore

Specificity

The only action named is "executing implementation plans" — a generic verb with no concrete capabilities (no mention of subagent dispatch, per-task review, fix loops, or final review). This matches the anchor 'Names the domain but actions are minimal or generic'; it is not a 3 because no 1-2 concrete skill capabilities are listed, and not a 1 because the domain is clearly named.

2 / 5

Completeness

Only the 'when' is present ("Use when executing implementation plans with independent tasks in the current session") with no 'what' — nothing states the skill dispatches subagents, reviews each task, or runs a final review. This exactly matches anchor 2 ('only when is present without what'); it is not 3 because that anchor requires a clear 'what' with a missing 'when' — the inverse of this description.

2 / 5

Trigger Term Quality

Relevant keywords exist — "executing implementation plans", "independent tasks", "current session" — and a user might naturally say "execute this plan". It stays at 3 rather than 4 because common variations and synonyms ("run the plan", "work through the tasks", "delegate tasks to subagents", "implementation plan") are missing; there is a single thin phrasing cluster.

3 / 5

Distinctiveness Conflict Risk

The qualifiers "independent tasks" and "in the current session" do distinguish it from a parallel-session plan-execution sibling, but "executing implementation plans" is a broad phrase that could overlap with plan-creation, code-review, or general execution skills. 'Somewhat specific but could still overlap with similar skills' is the best fit; not 4 because the distinction from the closely related executing-plans skill rests on one subtle qualifier.

3 / 5

Total

10

/

20

Passed

Validation

87%

Checks the skill against the spec for correct structure and formatting. All validation checks must pass before discovery and implementation can be scored.

Validation — 14 / 16 Passed

Validation for skill structure

CriteriaDescriptionResult

skill_md_line_count

SKILL.md is long (569 lines); consider splitting into references/ and linking

Warning

relative_links

Relative link issues: 4 missing, 1 suspicious

Warning

Total

14

/

16

Passed

Repository
openai/plugins
Reviewed

Table of Contents

Is this your skill?

If you maintain this skill, you can claim it as your own. Once claimed, you can manage eval scenarios, bundle related skills, attach documentation or rules, and ensure cross-agent compatibility.