CtrlK
BlogDocsLog inGet started
Tessl Logo

writing-skills

Use when creating new skills, editing existing skills, or verifying skills work before deployment

56

Quality

64%

Does it follow best practices?

Run evals on this skill

Adds up to 20 points to the overall score

View guide

SecuritybySnyk

Passed

No findings from the security scan

Fix and improve this skill with Tessl

tessl review fix ./eval/local/skills/benchmarks/dependency/superpowers/writing-skills/SKILL.md

The canonical home for this skill is writing-skills in obra/superpowers

SKILL.md
Quality
Evals
Security

Quality

Content

62%Weight 40%Scale 1-5

Reviews the quality of instructions and guidance provided to agents. Good implementation is clear, handles edge cases, and produces reliable results.

The body delivers an exceptionally clear, well-validated TDD workflow with concrete examples, templates, and checklists. Its weaknesses are significant verbosity with repeated content across sections, and a progressive-disclosure structure that inlines material it says to split while pointing to five bundle files that are not actually present.

Suggestions

Consolidate duplicated content: state the Iron Law and RED-GREEN-REFACTOR cycle once, merge the description guidance in "SKILL.md Structure" with the "Skill Discovery Optimization" section, and cut the body toward its own <500-word target by moving extended examples out.

Actually provide (or remove references to) the five cited bundle files — anthropic-best-practices.md, testing-skills-with-subagents.md, persuasion-principles.md, graphviz-conventions.dot, and render-graphs.js — since the core pressure-scenario methodology currently dead-ends at a missing file.

Inline a minimal fallback for writing pressure scenarios (the one operation the workflow depends on but delegates entirely to a missing file), so the skill remains actionable without the bundle.

DimensionReasoningScore

Conciseness

The body runs ~3,800 words — nearly 8x its own stated "<500 words" target for skills — with structural redundancy: the Iron Law appears in full twice ("The Iron Law" section and "The Bottom Line"), the RED-GREEN-REFACTOR cycle is explained three times (mapping table, dedicated section, checklist), and description-authoring rules are duplicated between "SKILL.md Structure" and "Skill Discovery Optimization". This is "noticeably verbose; several unnecessary explanations or padded sections" rather than the occasional tightening opportunity of anchor 3.

2 / 5

Actionability

Guidance is highly concrete: good/bad YAML description examples, a full SKILL.md structure template, exact commands ("wc -w skills/path/SKILL.md", "./render-graphs.js ../some-skill --combine"), a five-step micro-testing protocol, and per-item checklists. It falls short of fully executable because the core operation — how to actually write and run pressure scenarios — is deferred to testing-skills-with-subagents.md, a file that does not exist in the bundle.

4 / 5

Workflow Clarity

The RED → GREEN → REFACTOR sequence is explicit with validation checkpoints and feedback loops: "Run scenarios WITHOUT skill - document baseline behavior verbatim", "Run same scenarios WITH skill. Agent should now comply", "Re-test until bulletproof", plus a mandatory per-item creation checklist. This matches the top anchor: clear sequence, explicit validation steps, and checklists for a complex process.

5 / 5

Progressive Disclosure

References are one level deep and clearly signaled ("See anthropic-best-practices.md", "See [testing-skills-with-subagents.md](...)"), but the body inlines large sections its own rules say belong in separate files (extended SDO examples, testing methodology, code-example guidance) at ~3,800 words. More importantly, every referenced file — anthropic-best-practices.md, graphviz-conventions.dot, render-graphs.js, persuasion-principles.md, testing-skills-with-subagents.md — is missing from the bundle, so the navigation structure it advertises does not actually exist.

3 / 5

Total

14

/

20

Passed

Description

66%Weight 40%Scale 1-5

Based on the skill's description, can an agent find and select it at the right time? Clear, specific descriptions lead to better discovery.

The description has an explicit, specific "Use when..." trigger with concrete actions in a distinct niche. Its main weakness is the complete absence of a "what" statement — it describes only triggering conditions — and limited synonym coverage for natural search terms.

Suggestions

Add a brief "what" clause before the trigger, e.g., "Test-driven methodology for building skills: baseline-test agents, write minimal guidance, close loopholes. Use when creating new skills, editing existing skills, or verifying skills work before deployment."

Include common synonyms users would say, such as "writing", "authoring", or "documenting skills", to broaden natural keyword coverage.

Consider naming the core deliverable (a tested, deployable SKILL.md) so the description distinguishes this skill from general documentation or code-review skills.

DimensionReasoningScore

Specificity

The description names several concrete actions — "creating new skills, editing existing skills, or verifying skills work before deployment" — which are specific to the skill-authoring domain. It falls short of a 5 because it never states what the skill provides (e.g., a TDD-based testing methodology), listing only triggering activities.

4 / 5

Completeness

The "when" is explicit and concrete ("Use when creating new skills, editing existing skills, or verifying skills work before deployment"), but the description contains no "what" — it never says what the skill does or provides. This mirrors anchor 3 (clear one half, the other absent/implicit) rather than anchor 4, which requires both.

3 / 5

Trigger Term Quality

Phrases like "creating new skills", "editing existing skills", and "verifying skills work before deployment" match what a user would naturally say. Common synonyms such as "writing skills", "authoring", or "skill documentation" are missing, keeping it below a 5.

4 / 5

Distinctiveness Conflict Risk

The skill-authoring niche is distinct with concrete triggers (create/edit/verify skills), so it is unlikely to fire for unrelated skills. Minor overlap risk remains with other meta-skills about skill maintenance or documentation, keeping it below a 5.

4 / 5

Total

15

/

20

Passed

Validation

87%

Checks the skill against the spec for correct structure and formatting. All validation checks must pass before discovery and implementation can be scored.

Validation — 14 / 16 Passed

Validation for skill structure

CriteriaDescriptionResult

skill_md_line_count

SKILL.md is long (690 lines); consider splitting into references/ and linking

Warning

relative_links

Relative link issues: 1 missing

Warning

Total

14

/

16

Passed

Repository
rpamis/comet
Reviewed

Table of Contents

Is this your skill?

If you maintain this skill, you can claim it as your own. Once claimed, you can manage eval scenarios, bundle related skills, attach documentation or rules, and ensure cross-agent compatibility.