CtrlK
BlogDocsLog inGet started
Tessl Logo

planning-with-files

Use this by default for non-trivial multi-step work that needs persistent planning, progress tracking, or durable notes on disk. Trigger when a task will likely span multiple tool calls, research steps, verification loops, or enough context that the plan should not live only in transient chat memory.

59

Quality

70%

Does it follow best practices?

Run evals on this skill

Adds up to 20 points to the overall score

View guide

SecuritybySnyk

Passed

No findings from the security scan

Fix and improve this skill with Tessl

tessl review fix ./skills/planning-with-files/SKILL.md
SKILL.md
Quality
Evals
Security

Quality

Content

63%Weight 40%Scale 1-5

Reviews the quality of instructions and guidance provided to agents. Good implementation is clear, handles edge cases, and produces reliable results.

The body is strongly actionable — ready-to-use templates, a clear 3-file pattern, and per-phase commands — but it repeats its own trigger and rules across four sections, and its two advanced-content references point to files that do not exist in the bundle. Adding the missing reference files (or removing the dead links) and consolidating the duplicated sections would move this from good to excellent.

Suggestions

Create the referenced reference.md and examples.md (or remove the dead links) — the body currently points to two files that do not exist anywhere in the skill bundle.

Merge the "Use it proactively when", "When to Use This Pattern", "Critical Rules", and "Anti-Patterns" sections into a single trigger list and a single rules table; the same guidance is currently stated four times.

Add an explicit validation checkpoint before Loop 4 (e.g., "Re-read task_plan.md and confirm all phases are [x] and errors are resolved before producing the deliverable") to close the workflow's validation gap.

DimensionReasoningScore

Conciseness

The guidance itself is lean (tables, templates, short rules), but the same trigger conditions are restated four times — "Use it proactively when" (lines 13-19), "When to Use This Pattern" (lines 140-150), "Critical Rules" (lines 118-136), and the "Anti-Patterns" table (lines 152-160) largely repeat "plan first / store don't stuff / log errors" in different forms. This matches the "mostly efficient but could be tightened" anchor; not 4 because the duplication is substantive, not minor trimming.

3 / 5

Actionability

Copy-paste-ready templates for task_plan.md and notes.md (lines 70-116), a concrete 3-file table with update timing (lines 34-38), and specific tool commands per phase ("Read task_plan.md", "Edit task_plan.md", "Write notes.md", lines 52-64). Not 5 because the "Core Workflow" block (lines 42-47) is abstract summary rather than executable steps, and the bash-labeled blocks contain tool pseudocode rather than runnable commands.

4 / 5

Workflow Clarity

The 4-loop sequence (lines 43-47) plus "Before each major action / After each phase / When storing information" gives a clear, ordered workflow, and "Log All Errors" provides an error-recovery feedback loop. Not 5 because there is no validation checkpoint before delivery (e.g., re-reading the plan to confirm all phases are marked complete before producing the deliverable), leaving a minor validation gap at the anchor-4 level.

4 / 5

Progressive Disclosure

The structure is well organized (templates inline, advanced material split out with clearly labeled pointers: "See [reference.md](reference.md) for: Attention manipulation techniques..." and "See [examples.md](examples.md)"), but no references/, scripts/, or assets/ directories exist — both linked files are dead paths, so the advertised navigation does not work. This sits between anchor 3 (references present but not usable/well-organized) and anchor 4; the broken links keep it at 3 rather than the 4 the structure would otherwise earn.

3 / 5

Total

14

/

20

Passed

Description

76%Weight 40%Scale 1-5

Based on the skill's description, can an agent find and select it at the right time? Clear, specific descriptions lead to better discovery.

The description explicitly answers both what the skill does and when to use it, with good natural trigger keywords. Its main weakness is that the default-on, maximally broad trigger ("by default for non-trivial multi-step work") gives it near-universal overlap with every other multi-step skill, undermining distinctiveness.

Suggestions

Narrow the trigger from "by default for non-trivial multi-step work" to a bounded niche (e.g., tasks spanning multiple sessions, deliverable-heavy projects, or work that must be resumable), so it stops competing with every multi-step skill.

Add one concrete mechanism to the 'what' clause (e.g., 'maintains task_plan.md and notes.md as working memory') so the stated capabilities map to tangible artifacts.

DimensionReasoningScore

Specificity

Quotes: "persistent planning, progress tracking, or durable notes on disk" — three distinct, concrete capabilities are named, comparable to the anchor listing several specific actions with minor gaps. Not 5 because the actions are stated as abstract capabilities rather than concrete operations (no file pattern, artifacts, or mechanisms are named), and not 3 because three specific capabilities exceed the 1-2 actions of that anchor.

4 / 5

Completeness

Quotes: what — "needs persistent planning, progress tracking, or durable notes on disk"; when — "Trigger when a task will likely span multiple tool calls, research steps, verification loops, or enough context that the plan should not live only in transient chat memory". Both what and when are explicitly and clearly answered with concrete trigger phrases, matching the top anchor.

5 / 5

Trigger Term Quality

Quotes: "multi-step work", "planning", "progress tracking", "research steps", "verification loops", "tool calls" — good keyword coverage of phrases users would naturally say ("track progress", "plan this task"). Not 5 because common synonyms like "todo", "checklist", "roadmap", "break this down", or "milestones" are missing.

4 / 5

Distinctiveness Conflict Risk

Quotes: "Use this by default for non-trivial multi-step work" and "a task will likely span multiple tool calls" — the default-on framing over an extremely broad class of tasks creates high overlap risk with virtually any skill that involves multi-step work, matching the "very broad; high overlap risk with many similar skills" anchor. Not 3 because no domain boundary at all distinguishes it from other skills; not 1 because it is at least tied to planning/persistence rather than being fully generic.

2 / 5

Total

15

/

20

Passed

Validation

87%

Checks the skill against the spec for correct structure and formatting. All validation checks must pass before discovery and implementation can be scored.

Validation — 14 / 16 Passed

Validation for skill structure

CriteriaDescriptionResult

frontmatter_unknown_keys

Unknown frontmatter key(s) found; consider removing or moving to metadata

Warning

relative_links

Relative link issues: 2 missing

Warning

Total

14

/

16

Passed

Repository
Galaxy-Dawn/claude-scholar
Reviewed

Table of Contents

Is this your skill?

If you maintain this skill, you can claim it as your own. Once claimed, you can manage eval scenarios, bundle related skills, attach documentation or rules, and ensure cross-agent compatibility.