CtrlK
BlogDocsLog inGet started
Tessl Logo

constraint-driven-development

Establishes a project's quality bar as a written contract and stops agents quietly lowering it. Interviews the user on which dimensions matter, supplies sane default thresholds when they have no number in mind, records everything in CONSTRAINTS.md, and watches the diff for a weakened bar — new @ts-ignore or eslint-disable suppressions, skipped or deleted tests, assertions stripped out, unimplemented stubs, thresholds edited down. Use when no quality bar is written down, when the user says "set up constraints" or "define our standards", when the user wants dimensions they care about — accessibility, web performance, coverage — set up as enforced constraints, when an agent keeps silencing checks or skipping tests to get to green, when you need a coverage or performance threshold and don't know what number to pick, or when an agent writes more code than anyone will read.

75

Quality

94%

Does it follow best practices?

Run evals on this skill

Adds up to 20 points to the overall score

View guide
SecuritybySnyk

Passed

No findings from the security scan

SKILL.md
Quality
Evals
Security

Quality

Content

88%Weight 40%Scale 1-5

Reviews the quality of instructions and guidance provided to agents. Good implementation is clear, handles edge cases, and produces reliable results.

A highly actionable, well-sequenced process skill: concrete commands and templates at every step, explicit validation checkpoints, and a genuinely used one-level-deep reference implementation. The two weaknesses are moderate length for a SKILL.md — motivational and objection-handling prose that could be trimmed or split out — and some operational points stated more than once.

Suggestions

Trim or relocate the Common Rationalizations table and the Overview's motivational paragraphs — they are user-persuasion content, not operational guidance, and account for a meaningful share of the body's length.

State the diff-scoping/cost-placement rule once (Step 5) and reference it from Step 4's notes rather than repeating it in Step 4, Step 5, and Step 6.

Consider moving the full tool-install matrix (Step 4) into references/ alongside floor-guard.md so SKILL.md stays closer to an overview of the seven-step process.

DimensionReasoningScore

Conciseness

The body is consistently dense and operational — every table names tools, commands, and budgets, and it never explains concepts Claude already knows. It is not 5 because the Overview's motivational prose, the Common Rationalizations objection-handling table, and the diff-scoping/cost-placement rule stated three times (Step 4 note 3, Step 5 rules, Step 6) are trimmable; it is not 3 because nearly every section carries non-obvious, project-specific judgment rather than padding.

4 / 5

Actionability

Fully executable throughout: a copy-paste CONSTRAINTS.md template, a tool matrix with install and run commands ('tsc --noEmit', 'gitleaks detect --redact --no-banner', 'semgrep scan --config p/default'), a concrete package.json scripts block, and a phase-to-command lifecycle table. Commands are real and complete, matching the anchor for copy-paste-ready guidance covering common cases.

5 / 5

Workflow Clarity

Seven clearly sequenced steps (detect → interview → write → install → wire → guard → ratchet) with explicit validation checkpoints: the Verification checklist, the Red Flags self-check, the trial-run gate, and Q2's block-vs-warn failure semantics that define what happens when a check fails mid-task. Feedback-loop guidance is present, fitting the top anchor; there are no sequencing gaps to justify 4.

5 / 5

Progressive Disclosure

The one bundle file, references/floor-guard.md, is real, well-signaled twice ('A reference implementation of these five checks ships with this skill in [references/floor-guard.md]'), and exactly one level deep — the right split, since the guard script is implementation detail. It is not 5 because the ~310-line SKILL.md carries nearly everything inline (Sane Defaults, Common Rationalizations, the full tool-install matrix), reading more as a complete manual than an overview pointing to detailed materials.

4 / 5

Total

18

/

20

Passed

Description

96%Weight 40%Scale 1-5

Based on the skill's description, can an agent find and select it at the right time? Clear, specific descriptions lead to better discovery.

An exemplary description: third-person voice, concrete multi-action capability statement, an explicit and richly varied 'Use when' clause, and distinct, natural trigger phrasing. The only weakness is slight breadth into review-adjacent trigger territory, which creates minor conflict risk with sibling quality/review skills.

DimensionReasoningScore

Specificity

The description lists multiple concrete actions — 'Interviews the user on which dimensions matter, supplies sane default thresholds... records everything in CONSTRAINTS.md, and watches the diff for a weakened bar' — and enumerates the specific failure modes it detects ('new @ts-ignore or eslint-disable suppressions, skipped or deleted tests, assertions stripped out, unimplemented stubs, thresholds edited down'). This comprehensively matches the anchor-5 example's level of concrete, multi-action coverage; it is not score 4 because the action list goes beyond 'several' with no coverage gaps.

5 / 5

Completeness

It explicitly answers both what (establishes a written quality contract, interviews, records, and watches diffs for weakening) and when, with six concrete 'Use when' trigger clauses ('Use when no quality bar is written down, when the user says "set up constraints"...'). This matches the anchor-5 pattern of a clear what plus explicit, concrete trigger phrases; not 4 because the when-clause is fully explicit rather than improvable.

5 / 5

Trigger Term Quality

It includes natural phrases users would actually say — '"set up constraints"', '"define our standards"', 'accessibility, web performance, coverage', 'an agent keeps silencing checks or skipping tests to get to green' — plus synonyms (quality bar, constraints, standards, thresholds). This clearly matches the comprehensive-synonyms anchor; not 4, since the missing terms are trivial and both quoted user phrasings and the CONSTRAINTS.md filename are present.

5 / 5

Distinctiveness Conflict Risk

The niche — writing and enforcing a persistent quality-bar contract — is distinct, with triggers like 'set up constraints' and 'define our standards' unlikely to fire for the wrong skill. However, 'watches the diff for a weakened bar' and 'agent keeps silencing checks or skipping tests' sit adjacent to code-review skill territory, giving minor overlap risk with closely related skills, which fits anchor 4 rather than the minimal-conflict anchor 5.

4 / 5

Total

19

/

20

Passed

Validation

100%

Checks the skill against the spec for correct structure and formatting. All validation checks must pass before discovery and implementation can be scored.

Validation — 16 / 16 Passed

Validation for skill structure

No warnings or errors.

Repository
addyosmani/agent-skills
Reviewed

Table of Contents

Is this your skill?

If you maintain this skill, you can claim it as your own. Once claimed, you can manage eval scenarios, bundle related skills, attach documentation or rules, and ensure cross-agent compatibility.