CtrlK
BlogDocsLog inGet started
Tessl Logo

engineer-system-change

Evaluate and carry out non-trivial software-system changes from first principles. Use when assessing RFCs, issues, designs, features, refactors, migrations, dependency changes, or proposed fields, events, APIs, modules, and services whose need, consumers, system fit, validation, or rollback require scrutiny. Read the actual system, identify the concrete problem and named semantic consumers, choose the smallest sufficient solution, reject pseudo-requirements and speculative abstractions, and require evidence proportional to risk. Do not use for mechanical edits, source-code explanation, or a dedicated review of an already-complete diff.

76

Quality

94%

Does it follow best practices?

Run evals on this skill

Adds up to 20 points to the overall score

View guide

SecuritybySnyk

Passed

No findings from the security scan

SKILL.md
Quality
Evals
Security

Quality

Content

96%Weight 40%Scale 1-5

Reviews the quality of instructions and guidance provided to agents. Good implementation is clear, handles edge cases, and produces reliable results.

The body is an exceptionally lean, actionable, well-sequenced instruction skill with explicit verdicts and validation gates throughout. Progressive disclosure is the only minor gap: it is self-contained and well-structured but slightly long for a single file with no external references.

DimensionReasoningScore

Conciseness

Dense, lean prose that assumes Claude's competence with no padding or explanation of basic concepts; every line carries an instruction (e.g. 'Treat every proposed change as a hypothesis about a real system, not as an implementation checklist').

5 / 5

Actionability

Instruction-only skill with no code (appropriate per scoring notes) but highly concrete directives: 'Reproduce the baseline first. Encode it as a failing behavioral test when executable', 'Label material risk claims as VERIFIED, INFERENCE, or UNKNOWN', and a defined verdict vocabulary with explicit triggering conditions.

5 / 5

Workflow Clarity

A clearly sequenced six-gate workflow (Ground the Problem → Name Semantic Consumers → Choose the Smallest Sufficient Change → Map Consequences → Implement Only the Justified Slice → Prove the Result) with explicit validation checkpoints (return STOP/NEEDS_EVIDENCE, require observed evidence before implementation, prove the result against observed evidence) and recursive re-application of the gates.

5 / 5

Progressive Disclosure

No bundle files exist and the skill is a single self-contained SKILL.md with clear section headers and no nested references, which is well-organized; it scores just below 5 because the single file is fairly long and a couple of reference tables (e.g. the consumer ledger) could conceivably live in supporting files.

4 / 5

Total

19

/

20

Passed

Description

92%Weight 40%Scale 1-5

Based on the skill's description, can an agent find and select it at the right time? Clear, specific descriptions lead to better discovery.

The description is specific, complete, and distinctive, with concrete actions, explicit use-when triggers, and a sharp negative-scope clause. Its only minor weakness is trigger-term coverage, which is strong but lacks a few synonyms.

DimensionReasoningScore

Specificity

Lists multiple concrete actions across the whole lifecycle — 'assessing RFCs, issues, designs, features, refactors, migrations', 'identify the concrete problem and named semantic consumers', 'choose the smallest sufficient solution', 'reject pseudo-requirements' — with comprehensive coverage.

5 / 5

Completeness

Explicitly answers both 'what' ('Evaluate and carry out non-trivial software-system changes') and 'when' ('Use when assessing...') with concrete trigger phrases, and adds a negative 'Do not use for...' scope clause.

5 / 5

Trigger Term Quality

Strong natural vocabulary users would say ('RFCs, issues, designs, features, refactors, migrations, dependency changes') plus proposed artifacts (fields, events, APIs, modules, services); a few synonyms or casual phrasings are missing, so it sits just below the comprehensive anchor.

4 / 5

Distinctiveness Conflict Risk

Clear niche (non-trivial system changes whose need/consumers/rollback require scrutiny) and an explicit negative scope ('Do not use for mechanical edits, source-code explanation, or a dedicated review of an already-complete diff') sharply separates it from adjacent skills.

5 / 5

Total

19

/

20

Passed

Validation

100%

Checks the skill against the spec for correct structure and formatting. All validation checks must pass before discovery and implementation can be scored.

Validation16 / 16 Passed

Validation for skill structure

No warnings or errors.

Repository
bytedance/deer-flow
Reviewed

Table of Contents

Is this your skill?

If you maintain this skill, you can claim it as your own. Once claimed, you can manage eval scenarios, bundle related skills, attach documentation or rules, and ensure cross-agent compatibility.