CtrlK
BlogDocsLog inGet started
Tessl Logo

aw

Ships autonomous, end-to-end coding work — implement a feature or fix, all the way to a tested draft PR — from a single opt-in entry point. Detects the task tier (Micro / Lite / Full) and routes: Micro/Lite run single-pass in this context; Full hands off to the aw-planner → aw-executor agents. Use when the user asks to do a task "autonomously", "independently", "in isolation", "in a worktree", "end-to-end", "all the way to a PR", to "ship this", "land this", "take care of this", or "handle this without me" — or invokes `/aw` directly. Opt-in, not a wrapper on casual edits; the routing rule's exclusion list governs when to hold back. Triggers on "implement autonomously", "end-to-end", "in a worktree", "ship this", "/aw".

70

Quality

88%

Does it follow best practices?

Run evals on this skill

Adds up to 20 points to the overall score

View guide
SecuritybySnyk

Passed

No findings from the security scan

SKILL.md
Quality
Evals
Security

Quality

Content

81%Weight 40%Scale 1-5

Reviews the quality of instructions and guidance provided to agents. Good implementation is clear, handles edge cases, and produces reliable results.

A dense, highly actionable dispatcher policy: concrete commands and prompt templates, explicit gates, and a strict terminal contract make the workflow unambiguous. Its weak point is token efficiency — repeated rules and inline historical narration (including dated transcript analysis) pad the file, and some procedure detail could live in reference files rather than the main body.

Suggestions

State the capability-vs-name dispatch rule once (in Hard rules) and reference it from the Identity bullets, review-recovery, and dispatch-unavailable sections — the same rule is currently spelled out four times.

Move or delete the historical and time-sensitive narration (the ~/.claude/projects transcript analysis dated 2026-08-20 → 2026-09-25, and the account of the retired `aw` agent and the duplicated tier table) — it is drift-prone and adds no dispatch-time guidance.

Extract the full feature-pr-verifier dispatch procedure (four preconditions + prompt template) into a reference file such as references/verify-at-pr-open.md, keeping only the trigger and preconditions inline, to shorten the main body and make the referenced bundle structure real.

DimensionReasoningScore

Conciseness

Mostly load-bearing, skill-specific instruction, but noticeably loose in places: the capability-vs-name rule is stated four times (Identity bullets, "Test the capability, not the name", the "When sub-agent dispatch is unavailable" intro, and Hard rules), and inline history/time-sensitive narration like "In the local transcripts the restructure plan analysed (~/.claude/projects, 2026-08-20 → 2026-09-25)" and "It was duplicated between the dispatcher and SKILL.md while the dispatcher was an agent" adds drift-prone padding. Not 2: nothing explains generic concepts Claude already knows. Not 4: the repetition and historical narrative could be meaningfully tightened.

3 / 5

Actionability

Copy-paste-ready guidance throughout: `Skill("autonomous-workflow")`, exact memory.list/memory.write calls with scopes and tags, full `Task(subagent_type="aw-planner", ...)` and feature-pr-verifier prompt templates, `gh pr view "$PR" --json headRefOid,baseRefOid -q ...`, and verbatim MODE SELECTION and AW RUN COMPLETE output blocks. Not 4: the common cases, including the fallback paths, are all covered with executable commands.

5 / 5

Workflow Clarity

Clear sequence (Critical First Actions 1–3 → tier detection → routing table → follow-ups → terminal contract) with explicit validation checkpoints: "Clear the confidence(plan) ≥ 90% gate before writing any production code", "Phase 0 + Phase 2 stay mandatory in every tier", "Preconditions — all four, else record not run (<reason>)", and "Never report work you did not verify". Feedback loops are present (iterate-or-escalate below the gate, review recovery when the executor could not dispatch). Not 4: checkpoints are explicit at nearly every step, not just most.

5 / 5

Progressive Disclosure

Good structure with a deliberate, clearly signaled one-level split: "The tier table lives in exactly one place: autonomous-workflow/SKILL.md", "The lesson schema... live in rules/self-improvement-loop.md — read it rather than reasoning from the summary below". Not 5: no bundle files exist to verify the referenced paths (references/, rules/, and the sibling SKILL.md are absent from this bundle), and substantial procedure detail (the full feature-pr-verifier dispatch procedure and the review-recovery decision table) stays inline in an already long body. Not 3: references are well signaled with purpose statements and the split is intentional and navigable.

4 / 5

Total

17

/

20

Passed

Description

96%Weight 40%Scale 1-5

Based on the skill's description, can an agent find and select it at the right time? Clear, specific descriptions lead to better discovery.

An exemplary description: third-person, concrete, and explicit about both capability and trigger conditions, with an unusually rich set of natural trigger phrases plus the /aw invocation. The only weakness is that a few of the natural trigger phrases are broad enough that the opt-in guard clause has to do real work to prevent misfires.

DimensionReasoningScore

Specificity

Multiple concrete, third-person actions with comprehensive coverage: "Ships autonomous, end-to-end coding work — implement a feature or fix, all the way to a tested draft PR", "Detects the task tier (Micro / Lite / Full) and routes", "Full hands off to the aw-planner → aw-executor agents". Not 4: the core capability and its routing are fully described, not minor-gapped.

5 / 5

Completeness

Explicitly answers both what ("implement a feature or fix, all the way to a tested draft PR" with tier routing) and when ("Use when the user asks to do a task... or invokes /aw directly") with concrete trigger phrases. Not 4: both halves are explicit and concrete, not merely implied.

5 / 5

Trigger Term Quality

Comprehensive natural-phrase coverage including synonyms and the slash command: "autonomously", "independently", "in isolation", "in a worktree", "end-to-end", "all the way to a PR", "ship this", "land this", "take care of this", "handle this without me", "/aw". Not 4: no common variation is missing.

5 / 5

Distinctiveness Conflict Risk

The guard clause "Opt-in, not a wrapper on casual edits; the routing rule's exclusion list governs when to hold back" stakes out a clear niche, but everyday phrases like "take care of this" and "ship this" carry minor overlap risk with ordinary interactive coding requests — fits anchor 4 (mostly distinct, minor overlap) better than 5 (minimal conflict risk).

4 / 5

Total

19

/

20

Passed

Validation

75%

Checks the skill against the spec for correct structure and formatting. All validation checks must pass before discovery and implementation can be scored.

Validation — 12 / 16 Passed

Validation for skill structure

CriteriaDescriptionResult

metadata_field

'metadata' should map string keys to string values

Warning

frontmatter_unknown_keys

Unknown frontmatter key(s) found; consider removing or moving to metadata

Warning

relative_links

Relative link issues: 8 suspicious

Warning

referenced_paths_exist

Referenced path issues: 1 missing

Warning

Total

12

/

16

Passed

Repository
mthines/agent-skills
Reviewed

Table of Contents

Is this your skill?

If you maintain this skill, you can claim it as your own. Once claimed, you can manage eval scenarios, bundle related skills, attach documentation or rules, and ensure cross-agent compatibility.