CtrlK
BlogDocsLog inGet started
Tessl Logo

fusion-github-review-resolution

Resolves unresolved GitHub PR review threads end-to-end: evaluates whether each review comment is correct, applies a targeted fix when valid, replies with rationale when not, commits, and resolves the thread. USE FOR: unresolved review threads, PR review feedback, changes requested PRs, PR review URLs (#pullrequestreview-...), fix the review comments, close the open threads, address PR feedback. DO NOT USE FOR: summarizing feedback without code changes, creating new PRs, or read-only branches.

72

Quality

90%

Does it follow best practices?

Run evals on this skill

Adds up to 20 points to the overall score

View guide

SecuritybySnyk

Low

Low-risk findings worth noting

SKILL.md
Quality
Evals
Security

Quality

Content

81%Weight 40%Scale 1-5

Reviews the quality of instructions and guidance provided to agents. Good implementation is clear, handles edge cases, and produces reliable results.

A highly actionable, well-sequenced workflow with strong validation and error-recovery design for a mutation-capable batch process, backed by real, well-signaled bundle files. Its main weakness is conciseness: repeated rules and a duplicated trigger list add tokens without adding information.

Suggestions

State the reply-then-resolve ordering rule once (e.g. in step 6) and reference it from Safety & constraints instead of restating it four times across the document.

Replace the ~15-item "When to use" trigger enumeration with the 2-3 most common triggers and rely on the frontmatter description for the full list, cutting a substantial duplicated block.

Consider moving the GraphQL cost-awareness and token-budget guidance into a small assets/reference file, keeping only the mutation-cost rule and rate-limit warning inline in SKILL.md.

DimensionReasoningScore

Conciseness

Mostly efficient operational guidance, but noticeably loose in places: the "When to use" section re-enumerates ~15 trigger examples that duplicate the frontmatter description; the reply-then-resolve ordering rule is stated at least four times (steps 6.3's two-step mandate, "Never resolve a thread without a reply. Never post a reply without then resolving the thread.", "Post at most one reply attempt", and again under Safety & constraints); and "Do not claim checks passed unless commands were actually run" restates step 4's instruction. This fits the score-3 anchor ('mostly efficient but could be tightened') better than score 4's 'minor instances that could be trimmed' — the redundancy is recurring rather than incidental.

3 / 5

Actionability

Fully executable guidance: copy-paste-ready script invocations with real flags (e.g. `scripts/resolve-review-comments.sh --owner equinor --repo fusion-skills --pr 27 --review-id 3837647674 --apply --message "Addressed in <commit>: <what changed>."`), a concrete tooling map naming exact MCP tools and GraphQL asset files, and a specific `gh api graphql -f query=@assets/pull-request-review-threads.graphql` usage. Common cases (dry-run, apply, re-fetch on error) are covered with commands; this matches the score-5 anchor.

5 / 5

Workflow Clarity

The 9-step fetch → analyze → fix → validate → push → reply → resolve → verify pipeline has explicit validation checkpoints and feedback loops exactly where a batch mutation workflow needs them: targeted checks before required repo checks, dry-run-first with `--apply` gating, re-fetch thread state before retrying failed mutations, a final closure-verification step with baseline thread counts, and escalation for uncertain threads. This matches the score-5 anchor (clear sequence, explicit validation, error-recovery loops, and a bundled checklist for a complex process).

5 / 5

Progressive Disclosure

Structure is good: the body points to real, verified one-level-deep bundle files (all six `assets/*.graphql|md` and both `scripts/*.sh` exist), references are clearly signaled via the tooling map table and the "See each .graphql file in assets for complete mutation syntax" pointer, and the checklist is promoted as the working document. It falls short of the score-5 anchor because the ~220-line SKILL.md carries content that arguably belongs in references — the full trigger-phrase enumeration (duplicating the description) and the detailed GraphQL cost/token-budget guidance — making it a dense workflow document rather than a lean overview with well-split content.

4 / 5

Total

17

/

20

Passed

Description

100%Weight 40%Scale 1-5

Based on the skill's description, can an agent find and select it at the right time? Clear, specific descriptions lead to better discovery.

A strong description: third-person, specific, comprehensive trigger coverage including natural phrasings and a URL pattern, with explicit positive and negative use boundaries. It fully answers both 'what' and 'when' without padding.

DimensionReasoningScore

Specificity

The description lists multiple concrete actions — "evaluates whether each review comment is correct, applies a targeted fix when valid, replies with rationale when not, commits, and resolves the thread" — covering the full capability surface end-to-end. It is in third person ("Resolves..."), so no voice penalty applies; it matches the 'comprehensive coverage of specific concrete actions' anchor, not the score-4 anchor which expects minor gaps in coverage.

5 / 5

Completeness

It explicitly answers both what (evaluate, fix, reply, commit, resolve) and when ("USE FOR: unresolved review threads, PR review feedback...") with concrete trigger phrases, and additionally scopes exclusions ("DO NOT USE FOR: summarizing feedback..."). This is the score-5 anchor exactly; score 4 would require the 'when' to be less explicit or specific.

5 / 5

Trigger Term Quality

Trigger terms are comprehensive natural phrases and synonyms users would actually say: "unresolved review threads", "PR review feedback", "changes requested PRs", "fix the review comments", "close the open threads", "address PR feedback", plus a concrete URL pattern (#pullrequestreview-...). This matches the score-5 anchor (comprehensive coverage including synonyms); the score-4 anchor would require clearly missing natural terms, which is not the case.

5 / 5

Distinctiveness Conflict Risk

It occupies a clear niche (GitHub PR inline review-thread closure) with distinct triggers and an explicit negative scope that separates it from neighboring PR skills (summarizing, PR creation, read-only work). Minimal conflict risk with other skills; the score-4 anchor's 'minor overlap risk with closely related skills' is largely addressed by the DO NOT USE clause.

5 / 5

Total

20

/

20

Passed

Validation

93%

Checks the skill against the spec for correct structure and formatting. All validation checks must pass before discovery and implementation can be scored.

Validation — 15 / 16 Passed

Validation for skill structure

CriteriaDescriptionResult

metadata_field

'metadata' should map string keys to string values

Warning

Total

15

/

16

Passed

Repository
equinor/fusion-framework
Reviewed

Table of Contents

Is this your skill?

If you maintain this skill, you can claim it as your own. Once claimed, you can manage eval scenarios, bundle related skills, attach documentation or rules, and ensure cross-agent compatibility.