CtrlK
BlogDocsLog inGet started
Tessl Logo

using-spn

Use when addressing PR review threads on a GitHub PR — replying to review comments, resolving threads after fixes, looping through reviewer feedback, applying or counter-proposing suggestion blocks, sweeping outdated threads, checking PR mergeability, or discovering which PRs are open on a repository. Applies when the `spn` CLI is on PATH (`command -v spn`). Prefer spn over hand-rolled `gh api graphql` for review-thread work.

76

Quality

95%

Does it follow best practices?

Run evals on this skill

Adds up to 20 points to the overall score

View guide
SecuritybySnyk

Low

Low-risk findings worth noting

SKILL.md
Quality
Evals
Security

Quality

Content

88%Weight 40%Scale 1-5

Reviews the quality of instructions and guidance provided to agents. Good implementation is clear, handles edge cases, and produces reliable results.

A strong agent-facing CLI contract: executable examples, deterministic ordering, explicit error-recovery loops, and validation gates on every batch operation. The weaknesses are mild — some off-topic human-TUI/forks material and a monolithic single-file layout where reference files would slim the always-loaded body.

Suggestions

Cut the forks/CSV-mode section and the two `spn forks` Quick Reference rows down to a single pointer line to the `using-spn-forks` skill — the skill itself says fork work is a different use case, yet it still details the 24-column CSV header here.

Drop the "TUI parity" and "Spoon flag parallel" sections (human-facing tooling, not agent guidance), and collapse the BulkSkip.Reason table or the prose bullets above it — they state the same two rules twice within ten lines.

Move the stable reference material (error envelope shapes, exit-code catalog, rate-limit details, per-verb flag surfaces) into a references/ file (e.g. references/output-contract.md) and keep SKILL.md to the core loop, policies, and common mistakes.

DimensionReasoningScore

Conciseness

The bulk is non-derivable tool contract (exit codes, error envelopes, partial-failure dedup, per-verb flag surfaces) that earns its tokens, but there is trimmable material: the off-topic CSV-mode/forks section, the TUI parity and "Spoon flag parallel" sections about a human-facing tool, and the BulkSkip.Reason table duplicating the bullets immediately above it. Fits anchor 4 ("efficient; minor instances of over-explanation") more than 3, since the padding is a small fraction of the document.

4 / 5

Actionability

Every workflow ships copy-paste-ready bash: the core `spn threads next` loop with jq field extraction, the rate_limited retry, the partial-failure dedup retry, the dry-run gating pattern, and apply-suggestion/counter-proposal invocations. Fully executable and covering the common cases — matches the 5 anchor.

5 / 5

Workflow Clarity

The core loop is an explicit terminated sequence, and risky batch operations get explicit validation checkpoints: preview-then-act (`--dry-run` two-step and gating pattern), check `comment_posted` before retry, check `dryRun` before believing `isResolved`, and the mergeability gate before declaring done. Feedback loops for error recovery are present throughout, matching the 5 anchor; the destructive/batch cap does not apply because resolve-all is explicitly gated by dry-run previews.

5 / 5

Progressive Disclosure

Well-sectioned, one level deep, no nested references, with in-document pointers ("see Code context and verbose mode") and a clean hand-off to the separate `using-spn-forks` skill. However, at ~360 lines everything lives inline in SKILL.md — the error-contract detail, BulkSkip reference, and per-verb flag surface are reference-file candidates — which fits anchor 4 ("good structure; minor organization gaps") rather than 5 ("content appropriately split").

4 / 5

Total

18

/

20

Passed

Description

100%Weight 40%Scale 1-5

Based on the skill's description, can an agent find and select it at the right time? Clear, specific descriptions lead to better discovery.

The description is exemplary: it enumerates concrete capabilities, provides an explicit multi-trigger 'Use when' clause, names the tool and its availability check, and stakes out an unmistakable niche. Both the what and the when are concrete and user-phrased.

DimensionReasoningScore

Specificity

"replying to review comments, resolving threads after fixes, looping through reviewer feedback, applying or counter-proposing suggestion blocks, sweeping outdated threads, checking PR mergeability, or discovering which PRs are open" lists multiple specific concrete actions covering the tool's whole surface, matching the 5 anchor ("comprehensive coverage"). Not 4: there is no gap in the action coverage.

5 / 5

Completeness

Explicitly answers when ("Use when addressing PR review threads on a GitHub PR…") and what ("Prefer spn over hand-rolled `gh api graphql` for review-thread work") with concrete trigger phrases, exactly matching the 5 anchor. Not 4: the 'when' is not merely present but itemized and specific.

5 / 5

Trigger Term Quality

Natural user phrasings are covered with synonyms — "PR review threads", "review comments", "reviewer feedback", "suggestion blocks", "outdated threads", "mergeability", "open PRs" — plus the tool name and availability trigger ("command -v spn"). Not 4: no common way a user would phrase this task is missing.

5 / 5

Distinctiveness Conflict Risk

Clear niche (GitHub PR review threads via the spn CLI, gated on spn being on PATH) with distinct triggers and an explicit contrast to hand-rolled gh api graphql — minimal conflict risk with other skills. Not 4: it does not overlap meaningfully with any closely related skill's triggers.

5 / 5

Total

20

/

20

Passed

Validation

100%

Checks the skill against the spec for correct structure and formatting. All validation checks must pass before discovery and implementation can be scored.

Validation — 16 / 16 Passed

Validation for skill structure

No warnings or errors.

Repository
Raudbjorn/spoon
Reviewed

Table of Contents

Is this your skill?

If you maintain this skill, you can claim it as your own. Once claimed, you can manage eval scenarios, bundle related skills, attach documentation or rules, and ensure cross-agent compatibility.