CtrlK
BlogDocsLog inGet started
Tessl Logo

workthreads

SpecStory Workthreads - a weekly work-thread rollup across a team's repos from SpecStory coding histories (any agent - Claude Code, Codex, Cursor, Gemini, and more). It groups the window's sessions into threads of work per project and labels each new / open / recently closed, so a lead sees what shipped, what is still an open loop, and what was just started. Use when someone asks "what happened this week", "what is still open", "what did the team finish", "give me the weekly rollup", or wants a status report over a .specstory/history corpus.

74

Quality

93%

Does it follow best practices?

Run evals on this skill

Adds up to 20 points to the overall score

View guide
SecuritybySnyk

Low

Low-risk findings worth noting

SKILL.md
Quality
Evals
Security

Quality

Content

86%Weight 40%Scale 1-5

Reviews the quality of instructions and guidance provided to agents. Good implementation is clear, handles edge cases, and produces reliable results.

A tight, command-driven skill body: the engine does retrieval deterministically while the skill instructs only the judgment work, with verified CLI commands and a concrete deliverable shape. The main gaps are minor redundancy in the intro and the absence of explicit failure-mode checkpoints (empty digest, indexing errors) in the flow.

Suggestions

Trim the duplicated lifecycle description: the intro's "(new / open / recently closed)" preview repeats the full status definitions given in 'How the engine splits the work' — keep only the latter.

Add a brief failure-mode note to step 2, e.g. 'if the digest is empty, confirm the DB was indexed against the right --projects/--scan scope before concluding no work happened' — this adds the missing validation checkpoint for the batch indexing step.

Mention that --db defaults to ~/.specstory/workthreads.db so the index and threads commands are copy-paste runnable as-is for the common single-machine case.

DimensionReasoningScore

Conciseness

The body is lean and assumes competence ("you do the synthesis", no explanation of what transcripts or threads are conceptually), but the lifecycle is stated twice ("lines of work and their lifecycle (new / open / recently closed)" in the intro and again as full definitions in 'How the engine splits the work'), and the intro paragraph partially restates the frontmatter description. Not 5: those duplications could be trimmed; not 3: there is no genuinely unnecessary explanation or padding.

4 / 5

Actionability

Gives exact, verified commands — "node "${CLAUDE_SKILL_DIR}/scripts/workthreads.mjs" index --projects <parent-of-repos> --db <db>", "threads --db <db> --days 7" and "--days 7 --json" — plus a concrete output shape (sections a-e), a dated file convention, and "threads --out <file>". The flags match the actual CLI in scripts/workthreads.mjs (index/threads, --projects/--scan/--dir, --db, --days, --json, --out). Not 4: commands are copy-paste ready and cover the common cases; the only placeholders (<db>, <parent-of-repos>) are inherently user-specific paths.

5 / 5

Workflow Clarity

The default flow is a clearly numbered sequence (index → threads → write rollup → save to dated file) with a deliverable checklist (a-e) and an explicit evidence-citation requirement ("cite evidence refs (path:line) so each claim is checkable"). Not 5: there is no explicit checkpoint for failure modes — e.g., what to do if indexing errors, the digest comes back empty, or the DB was built against the wrong scope; the engine's own stderr/exit-code reporting is relied on implicitly. Not 3: the sequence is complete and the operations are non-destructive (read-only over the corpus, one report file written), and the 'checkable claims' instruction plus the in-progress caveat function as verification affordances.

4 / 5

Progressive Disclosure

SKILL.md is a concise overview; the implementation lives in a real, verified bundle (scripts/workthreads.mjs plus lib/ modules) referenced one level deep via the bash commands, with no nested references and well-organized sections (engine split, default flow, guided start, conventions). Not 4: nothing that belongs in a separate file is inlined — the digest format, statuses, and flags are all interface-level detail the body legitimately needs.

5 / 5

Total

18

/

20

Passed

Description

100%Weight 40%Scale 1-5

Based on the skill's description, can an agent find and select it at the right time? Clear, specific descriptions lead to better discovery.

The description follows the strongest exemplar pattern exactly: a concrete multi-action 'what' plus an explicit 'Use when...' clause with natural quoted trigger phrases and the corpus path. It is specific, complete, and clearly distinct from neighboring skills.

DimensionReasoningScore

Specificity

"a weekly work-thread rollup across a team's repos", "groups the window's sessions into threads of work per project", "labels each new / open / recently closed" — multiple concrete, specific actions with comprehensive coverage of the skill's behavior (rollup, grouping, lifecycle labeling). Not 4: coverage is comprehensive rather than having minor gaps; the actions named are exactly what the skill does.

5 / 5

Completeness

Explicitly answers both: what ("groups the window's sessions into threads of work per project and labels each new / open / recently closed") and when ("Use when someone asks 'what happened this week'...") with concrete trigger phrases. Not 4: the 'when' clause is explicit and enumerated, not merely present-but-imprecise.

5 / 5

Trigger Term Quality

"what happened this week", "what is still open", "what did the team finish", "give me the weekly rollup", "status report", ".specstory/history" — natural phrases a user would actually say, plus synonyms and the corpus path. Not 4: even common variations ("status report", "weekly rollup") are covered, matching the anchor-5 pattern of including synonyms and file extensions.

5 / 5

Distinctiveness Conflict Risk

"SpecStory coding histories", ".specstory/history corpus", "weekly work-thread rollup" define a clear niche with distinct triggers; minimal conflict with generic status/reporting skills. Not 4: although "status report" is a broad term, it is anchored to the SpecStory corpus, leaving minimal overlap risk.

5 / 5

Total

20

/

20

Passed

Validation

87%

Checks the skill against the spec for correct structure and formatting. All validation checks must pass before discovery and implementation can be scored.

Validation — 14 / 16 Passed

Validation for skill structure

CriteriaDescriptionResult

allowed_tools_field

'allowed-tools' contains unusual tool name(s)

Warning

frontmatter_unknown_keys

Unknown frontmatter key(s) found; consider removing or moving to metadata

Warning

Total

14

/

16

Passed

Repository
specstoryai/getspecstory
Reviewed

Table of Contents

Is this your skill?

If you maintain this skill, you can claim it as your own. Once claimed, you can manage eval scenarios, bundle related skills, attach documentation or rules, and ensure cross-agent compatibility.