CtrlK
BlogDocsLog inGet started
Tessl Logo

gate-check

Ready to advance between development phases? PASS/CONCERNS/NOT ASSESSED/FAIL with blockers and required artifacts. 'Can we move to production?'

57

Quality

72%

Does it follow best practices?

Run evals on this skill

Adds up to 20 points to the overall score

View guide
SecuritybySnyk

Passed

No findings from the security scan

Fix and improve this skill with Tessl

tessl review fix ./.claude/skills/gate-check/SKILL.md
SKILL.md
Quality
Evals
Security

Quality

Content

73%Weight 40%Scale 1-5

Reviews the quality of instructions and guidance provided to agents. Good implementation is clear, handles edge cases, and produces reliable results.

An exceptionally actionable and well-validated workflow with excellent reference-file delegation for the per-gate checklists. Its one real weakness is verbosity: design-rationale blockquotes and historical justification inflate the body well past what the executing model needs.

Suggestions

Move the why-it-must-be-this-way blockquotes (the 'zero required artifacts at rigor: minimal', 'performance.enforce is inert', and orphan-key rationale notes) into an authoring-notes or policy reference file, keeping only the operative rules inline — this is the main driver of the conciseness score of 2.

Extract the Section 8 follow-up remediation list (25+ items) into a references/follow-ups.md keyed by gate or missing artifact, and inline only the few relevant to the verdict.

Trim historical phrasing ('until now the verdict vocabulary had nowhere to put one', 'behavior unchanged from before this setting existed') to present-tense rules; the reader only needs the current behavior.

DimensionReasoningScore

Conciseness

The body is ~860 lines with several noticeably padded sections: long blockquote design-rationale digressions (e.g. 'Without that exclusion the Production → Polish gate had zero required artifacts...', 'A setting that works only when a rule is disregarded is not wired') and historical justifications ('until now the verdict vocabulary had nowhere to put one'). None of it explains concepts Claude already knows (it is all project-specific policy), so it is above a 1, but the repeated why-annotations are verbosity the runtime context does not need and belong in an authoring-notes reference.

2 / 5

Actionability

Fully executable throughout: exact commands ('bash .claude/scripts/artifact-check.sh --phase [source-phase]', 'mkdir -p production && printf ...'), ready-made AskUserQuestion prompts with option lists, concrete YAML templates for every project.yaml creation case, director gate IDs, and complete output templates. The specific examples cover the common cases (all three yaml-existence branches, per-tier panel tables).

5 / 5

Workflow Clarity

A clearly numbered sequence (Parse Arguments → Gate Definitions → Run Checks → Collaborative Assessment → Director Panel → Verdict → Stage Update → Next Steps) with exceptional validation checkpoints: the dedicated Chain-of-Verification section (5a) with explicit revise rules, the mandatory dual-write verification (6.3) with stop-on-divergence, confirm-before-write on stage changes, and precedence-ordered verdict rules ('FAIL, then CONCERNS, then NOT ASSESSED, then PASS') that remove inference. Feedback loops for error recovery are explicit throughout.

5 / 5

Progressive Disclosure

The six per-gate checklists are correctly split into real reference files (all six verified present in references/) with a clear table mapping each transition to its file and an explicit 'read only the row for the target phase' rule — a textbook one-level-deep split. It is not a 5 because the SKILL.md body itself still inlines large blocks that belong in references (the Section 2b tier-policy rationale and the 25-item Section 8 follow-up list), leaving the overview heavier than it needs to be.

4 / 5

Total

16

/

20

Passed

Description

61%Weight 40%Scale 1-5

Based on the skill's description, can an agent find and select it at the right time? Clear, specific descriptions lead to better discovery.

A compact, distinctive description with a strong quoted trigger phrase and concrete verdict vocabulary, but it names its outputs rather than its actions and lacks an explicit 'Use when...' trigger clause. It sits just above the midpoint of the scale.

Suggestions

Add an explicit trigger clause, e.g. 'Use when the user asks whether the project can advance to the next phase ("Can we move to production?", "are we ready for polish?")' — this would lift completeness from 3.

Name one or two of the concrete actions the skill performs (e.g. 'validates required artifacts and quality standards, reports blockers per phase gate') to strengthen specificity.

Include a couple of natural trigger variants ('phase gate', 'stage transition', 'ready for release') to round out keyword coverage.

DimensionReasoningScore

Specificity

The description names the domain ("advance between development phases") and concrete outputs ("PASS/CONCERNS/NOT ASSESSED/FAIL with blockers and required artifacts"), which is 1-2 concrete actions. It stops short of naming the checks performed (artifact checks, quality checks, director review), so it does not reach the 'several specific actions' of a 4 — and it is well above the vague 'names domain only' of a 2.

3 / 5

Completeness

The 'what' is clear (a verdict with blockers and required artifacts) but the 'when' is present only as a single quoted question with no 'Use when...' clause or equivalent explicit trigger guidance, which caps completeness at 3 per the judging guidelines. It is not a 2 because the quoted question does weakly establish the trigger, and not a 4 because the trigger condition is never stated as explicit guidance.

3 / 5

Trigger Term Quality

Good natural-keyword coverage: "advance between development phases", "production", and the directly quoted user question "'Can we move to production?'" is exactly what a user would say. A few natural variants are missing ("phase gate", "stage", "ready for"), keeping it below the comprehensive synonym/extension coverage of a 5, but it is clearly above the 'some relevant keywords, missing common variations' of a 3.

4 / 5

Distinctiveness Conflict Risk

The verdict vocabulary (PASS/CONCERNS/NOT ASSESSED/FAIL) and phase-advance framing form a fairly distinct niche with the quoted trigger. Minor overlap risk remains: "Can we move to production?" could naturally match a software-deployment or release skill, so it is below the 'clear niche, minimal conflict' of a 5 but above the 'could still overlap with similar skills' of a 3.

4 / 5

Total

14

/

20

Passed

Validation

81%

Checks the skill against the spec for correct structure and formatting. All validation checks must pass before discovery and implementation can be scored.

Validation — 13 / 16 Passed

Validation for skill structure

CriteriaDescriptionResult

skill_md_line_count

SKILL.md is long (892 lines); consider splitting into references/ and linking

Warning

allowed_tools_field

'allowed-tools' contains unusual tool name(s)

Warning

frontmatter_unknown_keys

Unknown frontmatter key(s) found; consider removing or moving to metadata

Warning

Total

13

/

16

Passed

Repository
Donchitos/Claude-Code-Game-Studios
Reviewed

Table of Contents

Is this your skill?

If you maintain this skill, you can claim it as your own. Once claimed, you can manage eval scenarios, bundle related skills, attach documentation or rules, and ensure cross-agent compatibility.