CtrlK
BlogDocsLog inGet started
Tessl Logo

autonomous-workflow

The phase-based machinery (0–7) behind the `aw` dispatcher — task intake through tested PR delivery in an isolated Git worktree, with optional companion skills for planning, quality gates, TDD, UX, code quality, docs, and CI verification. Companions never block; a skipped one is reported, never silent. NOT the entry point and not auto-triggered: a natural-language request to do work autonomously, end-to-end, in isolation, or in a worktree belongs to the `aw` skill, which detects the task tier and routes. Reach for this skill only to read or run the phase machinery directly, bypassing tier detection — invoke with /autonomous-workflow.

64

Quality

80%

Does it follow best practices?

Run evals on this skill

Adds up to 20 points to the overall score

View guide
SecuritybySnyk

Passed

No findings from the security scan

Fix and improve this skill with Tessl

tessl review fix ./skills/workflow/autonomous-workflow/SKILL.md
SKILL.md
Quality
Evals
Security

Quality

Content

70%Weight 40%Scale 1-5

Reviews the quality of instructions and guidance provided to agents. Good implementation is clear, handles edge cases, and produces reliable results.

The body delivers an unusually well-gated orchestration workflow — explicit phases, mandatory invariants, confidence thresholds, and real error-recovery loops — and the index-to-rules architecture is the right progressive-disclosure shape on paper. Its weaknesses are redundancy for a self-described thin index and a bundle whose referenced rule/template files are missing while most shipped reference files are unreachable from the index.

Suggestions

Consolidate the triplicated tier-decision content (the question walk, the tier table, and the 'Two readings this exists to rule out' prose) into one authoritative section, and deduplicate the `aw-tester` and `review-loop` entries that currently appear in multiple sections.

Move version-history prose ('since v3.23', dispatcher redesign rationale) into a dedicated rationale or changelog reference file so the operational index stays version-neutral and lean.

Either ship the referenced rules/, templates/, aw/, README.md, and CLAUDE.md files in the bundle or add links from the body to the 7 currently-orphaned files in references/ so every provided file is discoverable from SKILL.md.

DimensionReasoningScore

Conciseness

The body is dense and operational with no explanations of concepts Claude already knows, but at ~440 lines it contradicts its own 'thin index' claim: tier logic appears three times (the question walk, the tier table, and the 'Two readings this exists to rule out' prose), `aw-tester` is described in four separate sections, `review-loop` is listed twice in Related Skills, and version-sensitive prose ('since v3.23', "version '3.28.0'") sits outside any deprecated/old-patterns section. Not 2 because nothing is padded filler; not 4 because the repetition and design-rationale paragraphs are clearly trimmable.

3 / 5

Actionability

Provides copy-paste commands ('git clone https://github.com/mthines/agent-skills.git ... bash scripts/sync-symlinks.sh --aw', 'gw add fix/bug-name'), exact gates ('confidence(plan) ≥ 90%'), hard caps (3 Lite / 5 Full iterations), and a canonical MODE SELECTION output block. Not 5 because many per-phase actions defer to rules/*.md files rather than giving inline executable detail, so the common cases are not fully self-contained.

4 / 5

Workflow Clarity

The phases are sequenced 0–7 with a per-phase gate table, explicit invariants ('Phase 0 and Phase 2 are MANDATORY'), and genuine feedback loops: iteration cap triggers 'confidence(analysis)' then 'one-shot auto-replan or escalate to user', CI failure spawns 'ci-auto-fix' per failure, and gates allow 'OR user-approved stop'. Explicit validation checkpoints and error-recovery paths match the anchor-5 example; there are no missing-validation gaps to justify 4.

5 / 5

Progressive Disclosure

The index design is genuinely one-level-deep with well-signaled, purpose-described links ('Detailed procedures live in rules/*.md and load on demand'), but scored against the actual bundle the navigation breaks: ~48 of 50 relative links target files absent from the provided bundle (no rules/, templates/, aw/, README.md, or CLAUDE.md), and 7 of the 8 files in references/ are never linked from the body at all. Not 4 because missing referenced files and orphaned reference files are more than minor organization gaps; not 2 because the in-body structure and reference signaling are strong.

3 / 5

Total

15

/

20

Passed

Description

87%Weight 40%Scale 1-5

Based on the skill's description, can an agent find and select it at the right time? Clear, specific descriptions lead to better discovery.

The description is a strong two-skill disambiguation: it states concrete capabilities in third person, gives an explicit reach-for-it-only clause with an invocation command, and proactively routes the most likely natural-language triggers to the `aw` dispatcher. The only soft spot is that its natural trigger terms serve the sibling skill more than this one.

DimensionReasoningScore

Specificity

Names the domain ('phase-based machinery (0–7) behind the `aw` dispatcher') and several concrete capabilities — 'task intake through tested PR delivery in an isolated Git worktree' plus enumerated companion domains (planning, quality gates, TDD, UX, code quality, docs, CI verification). Falls short of 5 because the individual phase actions are not enumerated, and above 3 because more than 1–2 concrete actions are stated in third-person voice.

4 / 5

Completeness

Explicitly answers both: what ('task intake through tested PR delivery in an isolated Git worktree, with optional companion skills...') and when ('Reach for this skill only to read or run the phase machinery directly, bypassing tier detection — invoke with /autonomous-workflow'). The 'use when' equivalent is explicit and concrete, matching the anchor-5 pattern rather than the merely-adequate anchor 4.

5 / 5

Trigger Term Quality

Includes natural user phrases — 'do work autonomously', 'end-to-end', 'in isolation', 'worktree', and the explicit '/autonomous-workflow' invocation — with synonyms covered. Not 5 because the natural trigger phrases are mostly attributed to the sibling `aw` skill, leaving this skill's own direct-use triggers ('read or run the phase machinery directly') with thinner natural-term coverage.

4 / 5

Distinctiveness Conflict Risk

Actively disambiguates against the closest sibling: 'NOT the entry point and not auto-triggered: a natural-language request to do work autonomously... belongs to the `aw` skill'. This names and reroutes the competing trigger, giving a clear niche with minimal conflict risk — stronger than anchor 4's minor-overlap profile.

5 / 5

Total

18

/

20

Passed

Validation

75%

Checks the skill against the spec for correct structure and formatting. All validation checks must pass before discovery and implementation can be scored.

Validation — 12 / 16 Passed

Validation for skill structure

CriteriaDescriptionResult

metadata_field

'metadata' should map string keys to string values

Warning

frontmatter_unknown_keys

Unknown frontmatter key(s) found; consider removing or moving to metadata

Warning

relative_links

Relative link issues: 42 missing, 18 suspicious

Warning

referenced_paths_exist

Referenced path issues: 1 missing

Warning

Total

12

/

16

Passed

Repository
mthines/agent-skills
Reviewed

Table of Contents

Is this your skill?

If you maintain this skill, you can claim it as your own. Once claimed, you can manage eval scenarios, bundle related skills, attach documentation or rules, and ensure cross-agent compatibility.