CtrlK
BlogDocsLog inGet started
Tessl Logo

delivery-gate

Stop hook that blocks Claude from finishing until quality checks pass. Detects rationalization patterns (surface text heuristics), stale learning logs (filesystem mtime), and low disk space. Complements self-audit by mechanically enforcing learning capture habits. Use when Claude should be mechanically blocked from declaring work finished before quality checks and learning capture actually pass.

68

Quality

81%

Does it follow best practices?

Run evals on this skill

Adds up to 20 points to the overall score

View guide

SecuritybySnyk

Low

Low-risk findings worth noting

SKILL.md
Quality
Evals
Security

Quality

Content

75%Weight 40%Scale 1-5

Reviews the quality of instructions and guidance provided to agents. Good implementation is clear, handles edge cases, and produces reliable results.

An exemplary body in terms of token efficiency and internal structure — tables, terse examples, honest limitations, no padding. Its main defect is that the core script it instructs users to copy and edit (quality-gate.py) is not actually bundled, leaving the install and configuration steps unexecutable; a post-install verification step is also missing.

Suggestions

Bundle quality-gate.py in scripts/ (and reference it as scripts/quality-gate.py) so the install and configuration steps are actually executable.

Add a short verification step after install (e.g., echo a fake transcript to the hook or run it manually to confirm exit 0) to close the validation gap.

Reconcile naming: the skill is 'delivery-gate' but the script and install references are 'quality-gate.py' — pick one name to avoid confusion.

DimensionReasoningScore

Conciseness

Lean and dense: behavior is conveyed through tables (checks/mechanisms/on-hit, config variables), a copy-paste install block, and terse exit-code examples. It never explains concepts Claude already knows, and every section (checks, why, install, config, examples, limitations) earns its place — matching anchor 5 rather than anchor 4's 'minor instances that could be trimmed'.

5 / 5

Actionability

Concrete elements are strong — exact `cp` command, ready settings.json hook config, a config-variable table, and worked examples with exit codes and stderr text — but the entire skill depends on `quality-gate.py`, which is referenced ('cp quality-gate.py ~/.claude/scripts/', 'Edit quality-gate.py') yet absent from the bundle (no scripts/ directory). That is a key missing detail, so anchor 3 fits better than anchor 4's 'minor gaps'.

3 / 5

Workflow Clarity

The install workflow (copy script → register Stop hook in settings.json → create learning-library files → customize LIBS) is clearly sequenced, and the Examples section documents expected exit-code outcomes for each scenario, acting as informal checkpoints. It lacks an explicit post-install validation step (e.g., test-fire the hook or run it on a fake transcript), which keeps it at anchor 4 rather than anchor 5.

4 / 5

Progressive Disclosure

Well-organized sections with appropriately sized inline content and clearly signaled cross-references ('See Also' listing self-audit, verification-loop, gateguard), and the bundle contains no nested or buried references. However, the one hard file dependency, quality-gate.py, does not exist in the bundle, a real organization gap that anchor 5's 'easy navigation' would not have.

4 / 5

Total

16

/

20

Passed

Description

88%Weight 40%Scale 1-5

Based on the skill's description, can an agent find and select it at the right time? Clear, specific descriptions lead to better discovery.

A strong description: it states concrete, mechanism-annotated capabilities, includes an explicit and specific 'Use when' trigger clause, and clearly differentiates itself from the complementary self-audit skill. The only weaknesses are minor — a somewhat generic 'quality checks' phrase that could overlap with code-quality gate skills, and limited synonym coverage of trigger phrasings.

Suggestions

Add one or two synonym phrasings of the trigger (e.g., 'use when Claude tends to declare work done early' or 'session-completion gate') to broaden natural trigger coverage.

Sharpen 'quality checks pass' toward the mechanical checks it actually performs (mtime/regex/disk) to reduce co-trigger risk with code-review or verification skills.

DimensionReasoningScore

Specificity

Names all three concrete detection mechanisms — 'rationalization patterns (surface text heuristics)', 'stale learning logs (filesystem mtime)', and 'low disk space' — plus the blocking behavior. Coverage of the hook's capabilities is comprehensive and mechanism-annotated, matching the anchor-5 example's level of concrete detail; anchor 4 ('minor gaps in coverage') would understate it.

5 / 5

Completeness

Explicitly answers both parts: the 'what' (detects rationalization patterns, stale learning logs, low disk space; blocks finishing) and a literal 'Use when Claude should be mechanically blocked from declaring work finished before quality checks and learning capture actually pass' trigger clause. Not anchor 4, since the 'when' is fully explicit and specific rather than merely adequate.

5 / 5

Trigger Term Quality

Includes natural phrases a user would plausibly say: 'blocks Claude from finishing', 'quality checks pass', 'stale learning logs', 'low disk space', 'learning capture'. Falls short of anchor 5 because it lacks synonym variations users might phrase the need with (e.g., 'prevent Claude from declaring done', 'session gate', 'force learning capture'), but it clearly exceeds anchor 3's 'missing common variations' level.

4 / 5

Distinctiveness Conflict Risk

Occupies a clear niche (mechanical, deterministic Stop hook) and explicitly positions itself against 'self-audit' ('Complements self-audit by mechanically enforcing learning capture habits'). Minor overlap risk remains: generic 'quality checks pass' phrasing could co-trigger with code-review/verification-type skills, keeping it below anchor 5's 'minimal conflict risk'.

4 / 5

Total

18

/

20

Passed

Validation

100%

Checks the skill against the spec for correct structure and formatting. All validation checks must pass before discovery and implementation can be scored.

Validation — 16 / 16 Passed

Validation for skill structure

No warnings or errors.

Repository
affaan-m/ECC
Reviewed

Table of Contents

Is this your skill?

If you maintain this skill, you can claim it as your own. Once claimed, you can manage eval scenarios, bundle related skills, attach documentation or rules, and ensure cross-agent compatibility.