CtrlK
BlogDocsLog inGet started
Tessl Logo

sweep

Audit the current Fallow session for missed work, incomplete verification, stale documentation, companion drift, or cleanup before final completion.

67

Quality

80%

Does it follow best practices?

Run evals on this skill

Adds up to 20 points to the overall score

View guide

SecuritybySnyk

Passed

No findings from the security scan

SKILL.md
Quality
Evals
Security

Quality

Content

86%Weight 40%Scale 1-5

Reviews the quality of instructions and guidance provided to agents. Good implementation is clear, handles edge cases, and produces reliable results.

A lean, well-sequenced sweep workflow that is specific in its directives and explicit about verification, with a strong closing proof requirement. The only meaningful gap is the absence of executable commands and an explicit re-validation loop after fixes.

DimensionReasoningScore

Conciseness

Every step is a terse directive ("Re-read the user's full request, active plan, diff, and live pull-request state") with zero padding and no explanation of concepts Claude already knows; the four lesson filters each earn their place. This matches the lean anchor exactly, and there is nothing to trim without losing information.

5 / 5

Actionability

As an instruction-only skill, the guidance is concrete and specific (map acceptance criteria to evidence, verify branches/worktrees/artifacts, the four lesson filters, the strongest-form ordering), matching anchor 4. Not 5 because steps like the artifact-cleanliness check have no copy-paste-ready commands (e.g. git status --porcelain); not 3 because the directives are specific rather than vague.

4 / 5

Workflow Clarity

A clear 8-step sequence with explicit verification steps at positions 2, 4, and 6 plus a closing assertion ("Finding no obvious bug is not completion. Prove every requested outcome.") and a filter checklist in step 7 — anchor 4. Not 5 because there is no explicit feedback loop for error recovery (e.g. re-run verification until clean after fixing); not 3 because checkpoints are explicit, not merely implied.

4 / 5

Progressive Disclosure

The body is under 50 lines, self-contained, and needs no external references — no references/, scripts/, or assets/ directories exist and the body references none — and the numbered workflow is clearly organized. Per the simple-skill exception this matches the top anchor.

5 / 5

Total

18

/

20

Passed

Description

75%Weight 40%Scale 1-5

Based on the skill's description, can an agent find and select it at the right time? Clear, specific descriptions lead to better discovery.

A lean, specific description that names concrete audit targets and states when to use the skill, though the trigger phrasing leans on harness jargon. All dimensions land at 4: solid but each one step short of the top anchor.

Suggestions

Add an explicit user-voice trigger clause, e.g. "Use when the user asks to sweep, wrap up, or double-check the session before declaring a task complete."

Add a natural synonym for the niche terms, e.g. pairing "companion drift" with "companion repository out of sync", so users who describe the symptom in ordinary words still match.

Clarify "Fallow session" with a plain-language gloss (e.g. "the current Fallow (agent) session") so the description is not opaque outside the harness.

DimensionReasoningScore

Specificity

One concrete action verb ("Audit the current Fallow session") is paired with five specific audit targets ("missed work, incomplete verification, stale documentation, companion drift, or cleanup"), matching the several-specific-actions anchor with minor coverage gaps. It falls short of 5 because it lists a single action rather than multiple distinct concrete actions.

4 / 5

Completeness

Both "what" (audit for the five named failure modes) and "when" ("before final completion") are explicitly present, matching anchor 4. Not 5 because the when-guidance is a terse temporal qualifier rather than a concrete "Use when the user..." trigger phrase; not 3 because the when is explicitly stated, not merely implied.

4 / 5

Trigger Term Quality

Good natural keyword coverage ("audit", "missed work", "verification", "stale documentation", "cleanup", "before final completion") that users would plausibly say. Not 5 because terms like "companion drift" and "Fallow session" are tool-internal jargon rather than user-voice phrases, and no synonyms or extensions are offered.

4 / 5

Distinctiveness Conflict Risk

It carves a clear niche (end-of-task session sweep within a specific harness) distinct from general skills, but broadly applicable words like "audit", "cleanup", and "stale documentation" carry minor overlap risk with review/cleanup skills. Anchor 5 would require triggers with minimal conflict risk, which these shared words do not fully meet.

4 / 5

Total

16

/

20

Passed

Validation

100%

Checks the skill against the spec for correct structure and formatting. All validation checks must pass before discovery and implementation can be scored.

Validation — 16 / 16 Passed

Validation for skill structure

No warnings or errors.

Repository
fallow-rs/fallow
Reviewed

Table of Contents

Is this your skill?

If you maintain this skill, you can claim it as your own. Once claimed, you can manage eval scenarios, bundle related skills, attach documentation or rules, and ensure cross-agent compatibility.