CtrlK
BlogDocsLog inGet started
Tessl Logo

agent-response-style

Default response behavior for AI coding agents: professional, factual, neutral tone with calibrated critical evaluation. Baseline for every task: implementation, code review, debugging, planning, research, evaluation, Q&A.

56

Quality

63%

Does it follow best practices?

Run evals on this skill

Adds up to 20 points to the overall score

View guide

SecuritybySnyk

Passed

No findings from the security scan

Fix and improve this skill with Tessl

tessl review fix ./.claude/skills/agent-response-style/SKILL.md
SKILL.md
Quality
Evals
Security

Quality

Content

67%Weight 40%Scale 1-5

Reviews the quality of instructions and guidance provided to agents. Good implementation is clear, handles edge cases, and produces reliable results.

The body is a well-structured, genuinely actionable behavioral policy with clear sections, explicit workflows, and concrete number-bounded directives. Its main weakness is redundancy across the voice/phasing rules and a slightly long single-file body that repeats the design-vs-implementation distinction in three places.

Suggestions

State the design-phase (peer-review) vs implementation-phase (teacher) voice rule once in Project context and reference it elsewhere, instead of restating it in Baseline Behavior and again in the teacher section.

Trim the duplicated directives in the teacher section ("Make sure they understand the why" appears twice back-to-back) and merge the near-identical conciseness instructions from Baseline Behavior and Verbatim Directives.

Add one short before/after example of peer-review phrasing versus teacher phrasing to make "peer-review framing" concrete, or move the teacher/coach ritual into a separate reference file to slim the main body.

DimensionReasoningScore

Conciseness

The body is mostly efficient user-specific policy that assumes Claude's competence (it never explains basic concepts), but it has noticeable redundancy: the design-vs-implementation voice rule is stated in Project context, again in Baseline Behavior, and again opening the teacher section; "Make sure they understand the why" appears twice within consecutive lines; and "Be concise, direct" directives appear in two separate sections. This matches anchor 3 (some unnecessary explanation / could be tightened) better than anchor 2's pervasive padding.

3 / 5

Actionability

The guidance is concrete and number-bounded ("present 2 to 4 credible alternatives", "add 2 to 3 short related questions", the three-point problem/solution/broader-context list, "Quiz them with open-ended or multiple-choice questions with AskUserQuestion"), which is actionable for an instruction-only skill. Minor gaps keep it below 5: "peer-review framing" and "calibrated challenge" are named but never exemplified with a sample phrasing.

4 / 5

Workflow Clarity

The Suggested Response Workflow gives a clear ordered sequence (clarify the decision, compare 2-4 options, state fit/trade-offs/failure modes, name decisive evidence, optionally ask framing questions) that mirrors the Verbatim Directives. No validation checkpoints are required since the skill involves no destructive or batch operations, so the workflow-clarity cap at 3 does not apply; it is not a 5 because the workflow is only loosely bound to the two mandatory sections.

4 / 5

Progressive Disclosure

The skill is a single self-contained file with no bundle directories, and the body is organized into clearly labeled sections with working cross-references ("see *Project context*", "see *Act as a teacher and a personal coach*"). Structure is good, but at ~90 lines the body exceeds the under-50-line simple case and sections like the teacher/coach ritual or the humanizer policy could arguably live in separate reference files, which is a minor organization gap at anchor 4.

4 / 5

Total

15

/

20

Passed

Description

58%Weight 40%Scale 1-5

Based on the skill's description, can an agent find and select it at the right time? Clear, specific descriptions lead to better discovery.

The description clearly communicates what the skill does and declares an explicit always-on scope, but it describes response-style attributes rather than concrete actions and lacks natural user-facing trigger phrases. Its claim to be the baseline for every task also raises overlap risk with other broadly-scoped skills.

Suggestions

Rewrite the "when" clause as natural trigger phrases (e.g., "Use when responding to code reviews, design proposals, or implementation explanations, or whenever critique of a proposed design would help") so users' wording matches the trigger.

Add a few concrete behavioral actions to the "what" (e.g., "compares 2-4 alternatives, names trade-offs and failure modes, flags hidden assumptions") instead of only style adjectives.

Narrow the "Baseline for every task" claim to the task types where the calibrated-critique behavior actually matters most, reducing conflict with other general conduct skills.

DimensionReasoningScore

Specificity

The description names the domain ("Default response behavior for AI coding agents") and gives a couple of quasi-concrete attributes ("professional, factual, neutral tone with calibrated critical evaluation"), but these are style qualities rather than listed concrete actions. It goes beyond score 2's bare domain-naming, but does not list the several specific actions of score 4.

3 / 5

Completeness

The "what" is clear ("professional, factual, neutral tone with calibrated critical evaluation") and the "when" is present via equivalent explicit trigger guidance ("Baseline for every task: implementation, code review, debugging, planning, research, evaluation, Q&A"), so the missing-trigger cap at 3 does not apply. However, the "when" is phrased as a scope list rather than concrete user-side trigger phrases, keeping it below anchor 5.

4 / 5

Trigger Term Quality

Domain terms like "implementation, code review, debugging, planning, research, evaluation, Q&A" are relevant, but they function as scope declarations rather than natural phrases a user would say to invoke this skill, and no synonyms or variations are offered. Keyword coverage is partial rather than good, matching anchor 3 over anchor 4.

3 / 5

Distinctiveness Conflict Risk

The tone-plus-calibrated-critique niche is somewhat specific and distinguishable, but the "Baseline for every task" claim with an enumerated list of all task types is broad and risks overlapping other always-on style/conduct skills. It sits between the high-overlap anchor 2 and the mostly-distinct anchor 4, closer to 3.

3 / 5

Total

13

/

20

Passed

Validation

100%

Checks the skill against the spec for correct structure and formatting. All validation checks must pass before discovery and implementation can be scored.

Validation — 16 / 16 Passed

Validation for skill structure

No warnings or errors.

Repository
SRombauts/SQLiteCpp
Reviewed

Table of Contents

Is this your skill?

If you maintain this skill, you can claim it as your own. Once claimed, you can manage eval scenarios, bundle related skills, attach documentation or rules, and ensure cross-agent compatibility.