CtrlK
BlogDocsLog inGet started
Tessl Logo

executing-plans

Use when executing an implementation plan in the current session as the implementer yourself — your human partner chose inline execution, or no subagent tool is available

59

Quality

69%

Does it follow best practices?

Run evals on this skill

Adds up to 20 points to the overall score

View guide

SecuritybySnyk

Passed

No findings from the security scan

Fix and improve this skill with Tessl

tessl review fix ./skills/executing-plans/SKILL.md
SKILL.md
Quality
Evals
Security

Quality

Content

81%Weight 40%Scale 1-5

Reviews the quality of instructions and guidance provided to agents. Good implementation is clear, handles edge cases, and produces reliable results.

An exceptionally actionable, well-gated workflow: exact commands, real bundle scripts, explicit validation checkpoints and feedback loops, and exhaustive ledger formats. The main cost is token weight — the DOT graph in particular spends many lines restating prose that the surrounding sections already specify.

DimensionReasoningScore

Conciseness

The prose is dense and assumes Claude's competence (no basic-concept explanations), but the ~40-line DOT digraph restates every edge by repeating the full node labels — a large token cost duplicating what the Setup, Task Loop, and Final Review sections already specify — and the 13-row rationalizations table plus full example workflow add bulk that could be tightened. This lands between "mostly efficient with some unnecessary content" and "efficient with minor trimmable over-explanation".

3 / 5

Actionability

Every instruction is copy-paste executable: exact commands (`scripts/task-start PLAN_FILE N`, `task-done PLAN_FILE N BASE -- <test command>`, both of which exist in the bundle), exact ledger line formats (`Task <N>: complete (commits <base7>..<head7>, tests: <command> → <result>)`), exact ruling format, and a full worked example with realistic SHAs and command output.

5 / 5

Workflow Clarity

Setup → per-task loop → final review → finish is clearly sequenced with explicit validation checkpoints everywhere: the completion contract's evidence checklist, "watching it fail is a step, not a formality", three-outcome handling of every `Expected:` comparison with a named feedback loop (systematic-debugging), and task-done refusing to record on a failing run.

5 / 5

Progressive Disclosure

Structure is good: the heavy lifting is correctly delegated one level deep to real, clearly signaled destinations — this bundle's `scripts/task-start` and `scripts/task-done` (both present on disk) and sibling-skill files (`../subagent-driven-development/scripts/sdd-workspace`, `../requesting-code-review/code-reviewer.md`). Minor gaps: everything else lives inline in one ~370-line SKILL.md, and the common-rationalizations table and example workflow could sit in a references file if the skill grows.

4 / 5

Total

17

/

20

Passed

Description

57%Weight 40%Scale 1-5

Based on the skill's description, can an agent find and select it at the right time? Clear, specific descriptions lead to better discovery.

The description has an explicit and well-conditioned "Use when" clause and good distinctiveness against its sibling skill, but the "what" is a single thin action and it misses natural trigger variations. The second-person voice ("yourself", "your human partner") triggers the rubric's specificity penalty.

Suggestions

Rewrite in third person (e.g., "Executes an implementation plan inline in the current session") to avoid the second-person specificity penalty from "yourself" / "your human partner".

Add natural trigger variations users would actually say, such as "implement the plan yourself", "run the plan in this session", or "execute the plan without subagents".

Enrich the "what" with the skill's concrete components — per-task TDD gating, a persisted ledger, one whole-branch final review — to raise specificity and completeness.

DimensionReasoningScore

Specificity

The description names one concrete action — "executing an implementation plan in the current session as the implementer yourself" — which fits the anchor of naming the domain with minimal actions, but it uses second-person voice ("yourself", "your human partner"), which the rubric penalizes by reducing specificity by 1 from 3. It lists no multiple specific capabilities (ledger, TDD gating, final review) that the skill actually performs.

2 / 5

Completeness

Both what ("executing an implementation plan in the current session as the implementer yourself") and when ("your human partner chose inline execution, or no subagent tool is available") are explicitly stated with a "Use when" clause. The "when" is highly specific with two concrete conditions, but the "what" is a single thin clause that undersells the skill's actual machinery.

4 / 5

Trigger Term Quality

Relevant keywords are present — "executing an implementation plan", "inline execution", "no subagent tool is available" — but common natural variations are missing ("implement the plan yourself", "run the plan in this session", "do it yourself"). This sits between "some relevant keywords but missing variations" and "good keyword coverage".

3 / 5

Distinctiveness Conflict Risk

The triggers ("inline execution", "no subagent tool") carve a clear niche and explicitly contrast with subagent-driven execution, keeping conflict risk low. Minor overlap risk remains with superpowers:subagent-driven-development, whose territory (executing a written plan) this description also names.

4 / 5

Total

13

/

20

Passed

Validation

93%

Checks the skill against the spec for correct structure and formatting. All validation checks must pass before discovery and implementation can be scored.

Validation — 15 / 16 Passed

Validation for skill structure

CriteriaDescriptionResult

relative_links

Relative link issues: 1 suspicious

Warning

Total

15

/

16

Passed

Repository
obra/superpowers
Reviewed

Table of Contents

Is this your skill?

If you maintain this skill, you can claim it as your own. Once claimed, you can manage eval scenarios, bundle related skills, attach documentation or rules, and ensure cross-agent compatibility.