CtrlK
BlogDocsLog inGet started
Tessl Logo

white-bear

Use when the user asks to "check white bear", "audit prohibitions", "find negative framing", or invokes /white-bear. Read-only audit: unnecessary competing-target mentions in LLM-facing prose.

62

Quality

78%

Does it follow best practices?

Run evals on this skill

Adds up to 20 points to the overall score

View guide
SecuritybySnyk

Passed

No findings from the security scan

Fix and improve this skill with Tessl

tessl review fix ./epistemic-cooperative/skills/white-bear/SKILL.md
SKILL.md
Quality
Evals
Security

Quality

Content

67%Weight 40%Scale 1-5

Reviews the quality of instructions and guidance provided to agents. Good implementation is clear, handles edge cases, and produces reliable results.

A well-structured, actionable audit procedure with concrete decision criteria and a complete output contract. Its main weakness is token efficiency: psychological rationale, hedged epistemics, and sections like Distinction/Confidence add length without adding instructions.

Suggestions

Trim the background rationale (white-bear effect explanation, ironic-process discussion in "What to evaluate") to one sentence each — Claude knows the effect; only the operative decision rules are new information.

Condense or merge the Distinction and Confidence sections into a few lines; phrases like "each maintains its own confidence curve" and "the audit illuminates the decision and leaves the judgment with the author" restate the Purpose section without adding guidance.

Move the per-form tests and load-bearing-boundary definitions (the longest stretch of "What to evaluate") into a single reference file, keeping only the constitutive rewrite test and severity table in SKILL.md, to bring the body closer to the token budget.

DimensionReasoningScore

Conciseness

The judgment criteria (rewrite test, per-form tests, boundary rules, severity table) are load-bearing, but the prose also explains background Claude already knows (the white bear effect, "the human ironic-process effect does not transfer mechanistically to language models") and carries padded qualifiers ("each maintains its own confidence curve", "the audit illuminates the decision and leaves the judgment with the author"). This is the mostly-efficient-with-unnecessary-explanation anchor, not anchor 4 where only minor trims would be needed.

3 / 5

Actionability

For an instruction-only skill the guidance is mostly executable: a constitutive rewrite test, per-form checklists, a severity calibration table, and a fully specified output JSON including zero-findings handling. It falls short of anchor 5 because several decision points remain soft judgment calls ("usually `low` for human triage", "judgment-dependent rewrite preference") rather than checkable rules.

4 / 5

Workflow Clarity

Sections sequence cleanly from Inputs to Scope to What-to-evaluate to Output to Self-application, with checkpoints (severity calibration, mandatory summary emission, zero-findings rule). The skill is read-only so no destructive-validation cap applies; it is not a 5 because there is no explicit feedback or error-recovery loop for ambiguous findings beyond "treat as low".

4 / 5

Progressive Disclosure

No bundle files exist; the body is a self-contained, well-headed document with no nested or buried references, matching the good-structure anchor. It is not a 5 because the skill exceeds the ~50-line simple-skill exemption and the long definitional block in "What to evaluate" is inline content that could plausibly live in a one-level-deep reference file.

4 / 5

Total

15

/

20

Passed

Description

82%Weight 40%Scale 1-5

Based on the skill's description, can an agent find and select it at the right time? Clear, specific descriptions lead to better discovery.

A strong description: explicit what and when, concrete trigger phrases, and a clearly distinct niche. The main gap is that only one action (the audit itself) is stated rather than the several concrete capabilities the skill body actually covers.

DimensionReasoningScore

Specificity

The description names the domain ("LLM-facing prose") and one concrete action ("Read-only audit: unnecessary competing-target mentions") but does not enumerate the multiple audit forms or actions, matching the anchor for domain plus 1-2 concrete actions. It is not a 4 because no several specific actions are listed, and not a 2 because the action stated is concrete rather than generic.

3 / 5

Completeness

It explicitly answers both questions: an explicit "Use when..." clause with concrete trigger phrases, and an explicit what ("Read-only audit: unnecessary competing-target mentions in LLM-facing prose"), matching the anchor-5 example pattern exactly.

5 / 5

Trigger Term Quality

Concrete quoted triggers ("check white bear", "audit prohibitions", "find negative framing", "/white-bear") with synonym variation give good natural-term coverage, though a few phrasings a user might naturally say (e.g., asking to remove "don't do X" phrasing) are missing — the anchor-4 case rather than comprehensive anchor-5 coverage.

4 / 5

Distinctiveness Conflict Risk

A clear niche (semantic audit for one specific authoring principle) with distinct, non-generic trigger phrases; minimal overlap risk with other skills, matching the anchor-5 case.

5 / 5

Total

17

/

20

Passed

Validation

87%

Checks the skill against the spec for correct structure and formatting. All validation checks must pass before discovery and implementation can be scored.

Validation — 14 / 16 Passed

Validation for skill structure

CriteriaDescriptionResult

allowed_tools_field

'allowed-tools' contains unusual tool name(s)

Warning

frontmatter_unknown_keys

Unknown frontmatter key(s) found; consider removing or moving to metadata

Warning

Total

14

/

16

Passed

Repository
jongwony/epistemic-protocols
Reviewed

Table of Contents

Is this your skill?

If you maintain this skill, you can claim it as your own. Once claimed, you can manage eval scenarios, bundle related skills, attach documentation or rules, and ensure cross-agent compatibility.