CtrlK
BlogDocsLog inGet started
Tessl Logo

behavioral-design

Applies behavioral science to make a desired behavior the easy default — for team adoption of tools, practices, and process changes as much as for product flows. Diagnoses why a behavior is not happening (COM-B, Fogg B=MAP), designs ethical interventions (EAST, choice architecture, defaults, friction, System 1 / System 2 fit), audits rollout plans, and builds behavior-change workshops and presentations. Never recommends a dark pattern or sludge. Use when the question is "why don't people do X" or "how do we get people to adopt X", not "is this component usable". Triggers on "behavioral design", "behavior change", "nudge", "drive adoption", "why aren't people using", "make it easy to adopt", "system 1 / system 2", "thinking fast and slow", "/behavioral-design".

68

Quality

85%

Does it follow best practices?

Run evals on this skill

Adds up to 20 points to the overall score

View guide
SecuritybySnyk

Passed

No findings from the security scan

SKILL.md
Quality
Evals
Security

Quality

Content

67%Weight 40%Scale 1-5

Reviews the quality of instructions and guidance provided to agents. Good implementation is clear, handles edge cases, and produces reliable results.

A superbly written control document — lean, well-sequenced, with a thoughtful mode-dispatch and on-demand reading table — that is undermined by its bundle: only references/frameworks.md ships, while all rules/ and templates/ files the workflow depends on are absent. As delivered, the skill can route and frame a request but cannot execute its diagnosis, ethics gate, or emission steps.

Suggestions

Ship the referenced bundle files (rules/diagnosis.md, rules/interventions.md, rules/dual-process.md, rules/ethics.md, rules/workshop.md, templates/intervention-plan.md, templates/workshop.md) or inline their essential content into SKILL.md — currently 7 of 8 referenced paths are dangling, which is the single largest quality defect.

Inline the three-question ethics test (or its essence) in Step 4 so the mandatory gate remains executable even if rules/ethics.md fails to load, mirroring the self-contained fallback pattern already used for missing composed skills.

Add a one-line fallback for missing rule files, e.g. proceed with the Hard Rules as the minimum viable rule set and report which files were skipped, matching the existing "A missing skill never blocks" convention.

DimensionReasoningScore

Conciseness

The body is lean and assumes Claude's competence: frameworks are named ("COM-B, Fogg B=MAP") without tutorials, the ✗/✓ rewrite example in Step 1 teaches the rule in two lines, and the mode and reading tables replace prose entirely. No section explains concepts Claude already knows, so nothing is trimmable without losing function — anchor 4's "minor instances of over-explanation" does not apply.

5 / 5

Actionability

The inline process scaffolding is concrete (the who/what/when/measured rewrite formula with a worked example, the batched four-item context checklist with `unknown` handling, testable Hard Rules like "Every intervention has a metric and a baseline-or-`unknown`"), but the operational core is deferred to files absent from the bundle: Step 3 says "Follow the loaded rule files", Step 4 gates output on "the three-question test in rules/ethics.md", and Step 5 fills "the mode's template" — and of the eight referenced paths only references/frameworks.md exists. This matches the concrete-but-incomplete anchor; it is not 4 because the gaps are the levers, ethics test, and templates themselves.

3 / 5

Workflow Clarity

Steps 1–5 are clearly sequenced with real checkpoints: the attitude-vs-behavior check in Step 1 ("ask for the action once"), unknown-baseline recovery in Step 2 ("proceed with `unknown`s… never stall"), the Step 4 ethics gate ("Any intervention that fails is dropped from the output and replaced by its honest alternative"), and the Hard Rules acting as an emission checklist. It is not 5 because the central validation checkpoint's content — the three-question ethics test — lives only in the missing rules/ethics.md, so the gate is named but not executable as written.

4 / 5

Progressive Disclosure

Scored against the actual bundle structure: the body designs an exemplary one-level-deep, load-on-demand scheme (the "Required Reading by Mode" table, "Load on demand — never all up-front", and references/frameworks.md § Evidence grading — which does exist), but only 1 of the 8 referenced paths is present; the five rules/*.md and two templates/*.md files are missing, so navigation to the core content is impossible in practice. The majority of paths being dangling pulls this between anchor 2 and anchor 3, landing at 2.

2 / 5

Total

14

/

20

Passed

Description

100%Weight 40%Scale 1-5

Based on the skill's description, can an agent find and select it at the right time? Clear, specific descriptions lead to better discovery.

An exemplary description: concrete third-person capabilities with named frameworks, an explicit Use-when clause with natural trigger phrases and a disambiguating negative boundary. Every token earns its place; nothing is padded or over-claimed.

DimensionReasoningScore

Specificity

The description lists multiple concrete actions — "Diagnoses why a behavior is not happening (COM-B, Fogg B=MAP)", "designs ethical interventions (EAST, choice architecture, defaults, friction, System 1 / System 2 fit)", "audits rollout plans", "builds behavior-change workshops and presentations" — each anchored to named frameworks, matching the comprehensive-coverage anchor. It stays in third person throughout, and even the exclusion ("Never recommends a dark pattern or sludge") is a concrete commitment rather than filler.

5 / 5

Completeness

Both questions are answered explicitly: the "what" is the enumerated capability set (diagnose, design, audit, workshop), and the "when" is a literal "Use when" clause with concrete trigger phrases plus a negative boundary ("not 'is this component usable'") that sharpens when-not-to-use. This matches the top anchor's pattern of what + when with concrete triggers.

5 / 5

Trigger Term Quality

Trigger coverage is comprehensive and natural: "why don't people do X", "how do we get people to adopt X", "behavior change", "nudge", "drive adoption", "why aren't people using", "make it easy to adopt", "system 1 / system 2", "thinking fast and slow", and the explicit "/behavioral-design" command. These are phrases a user would genuinely say, including synonyms and the slash-invocation form.

5 / 5

Distinctiveness Conflict Risk

The behavioral-science/adoption niche is distinct, and the description actively disambiguates from the nearest neighbor (usability review) via "not 'is this component usable'". Trigger phrases like "nudge" and "system 1 / system 2" are owned by this domain, so wrong-skill triggering risk is minimal.

5 / 5

Total

20

/

20

Passed

Validation

81%

Checks the skill against the spec for correct structure and formatting. All validation checks must pass before discovery and implementation can be scored.

Validation — 13 / 16 Passed

Validation for skill structure

CriteriaDescriptionResult

metadata_field

'metadata' should map string keys to string values

Warning

frontmatter_unknown_keys

Unknown frontmatter key(s) found; consider removing or moving to metadata

Warning

relative_links

Relative link issues: 19 missing

Warning

Total

13

/

16

Passed

Repository
mthines/agent-skills
Reviewed

Table of Contents

Is this your skill?

If you maintain this skill, you can claim it as your own. Once claimed, you can manage eval scenarios, bundle related skills, attach documentation or rules, and ensure cross-agent compatibility.