CtrlK
BlogDocsLog inGet started
Tessl Logo

pr-review-merge

Autonomous, safety-first pull request review-and-merge pass over a repository with many open PRs, on any tech stack (Node, PHP, Go, Python, Rust, Ruby, Java/Kotlin, Android, monorepos). Reviews each PR's full diff, classifies it CLEAN / FIXABLE / BLOCKED, fixes small issues, merges safe PRs one at a time, verifies the default branch still builds and passes tests, and produces a final report. Use this skill whenever the user mentions a backlog of pull requests, asks to review and merge PRs, clear the PR queue, "merge what's safe", check that nothing breaks after merging, or hands PR handling to an agent while they are away, even if they don't say "skill" or name a specific stack. Supports "review only" / "dry run" mode.

78

Quality

98%

Does it follow best practices?

Run evals on this skill

Adds up to 20 points to the overall score

View guide

SecuritybySnyk

Low

Low-risk findings worth noting

SKILL.md
Quality
Evals
Security

Quality

Content

100%Weight 40%Scale 1-5

Reviews the quality of instructions and guidance provided to agents. Good implementation is clear, handles edge cases, and produces reliable results.

Excellent content: dense, decision-oriented, and complete for an unattended batch operation — baseline capture, per-merge verification, a revert circuit breaker, and honest reporting requirements are all explicit. Progressive disclosure is exemplary, with stack-specific and template detail correctly externalized into two real, one-level reference files.

DimensionReasoningScore

Conciseness

The body is lean and imperative throughout — it assumes competence (never explains what a PR, lint, or CI is) and every line instructs a decision or action ('Read the full diff, not just file names', 'Merge one PR at a time', 'revert with a revert commit (no history rewriting)'). Even the brief rationale lines ('They exist because this runs unattended') earn their place by changing decisions, so nothing qualifies as the trimmable over-explanation of the level-4 anchor.

5 / 5

Actionability

Guidance is concrete and executable at every step: exact files to inspect (CONTRIBUTING, `.github/workflows/*.yml`, lockfiles), concrete tooling to use (`gh`/`glab`, the repo's Gradle/Maven wrapper), explicit classification criteria, and unambiguous rules for fixes, merges, and reporting. Per the rubric's instruction-skill note, the absence of code blocks is not penalized; the specificity matches the copy-paste-readiness of the level-5 anchor.

5 / 5

Workflow Clarity

Steps 0–5 are clearly sequenced with explicit validation checkpoints and feedback loops: run a baseline before touching anything, re-run checks after each fix and each merge, revert-with-a-revert-commit and stop merging on regression (circuit breaker), flaky-test re-run policy, and a final verification pass compared against the baseline. This is a batch operation and every validation the rubric demands is present, matching the level-5 anchor.

5 / 5

Progressive Disclosure

The body is a well-organized overview, and the genuinely detailed material is split into exactly the right two one-level-deep references, both real and clearly signaled at the point of use: 'Read `references/stack-discovery.md` for a lookup of where each stack keeps its build/test/lint config' and 'Use `references/report-template.md` for the exact structure.' Neither file nests further references, matching the level-5 anchor.

5 / 5

Total

20

/

20

Passed

Description

96%Weight 40%Scale 1-5

Based on the skill's description, can an agent find and select it at the right time? Clear, specific descriptions lead to better discovery.

A strong description: comprehensive, concrete, third-person, and equipped with an explicit and well-phrased trigger clause covering natural user language and a dry-run mode. The only weakness is minor overlap risk with generic code-review skills on the shared 'review' trigger terms.

Suggestions

Narrow the review-related triggers toward the backlog/merge intent (e.g., emphasize 'many open PRs' or 'backlog' as the primary trigger) to reduce overlap with a single-PR code-review skill.

Consider dropping or scoping the 'review only' / 'dry run' mention to 'backlog review only' so a simple 'review my PR' request does not activate this skill.

DimensionReasoningScore

Specificity

Lists multiple concrete actions — 'Reviews each PR's full diff, classifies it CLEAN / FIXABLE / BLOCKED, fixes small issues, merges safe PRs one at a time, verifies the default branch still builds and passes tests, and produces a final report' — with comprehensive coverage of the whole review-and-merge workflow. It is third-person ('Reviews', 'classifies', 'merges') and contains no vague filler, so it is not the level-4 anchor (which implies coverage gaps).

5 / 5

Completeness

Explicitly answers both questions: the 'what' is the detailed action list and classification model, and the 'when' is an explicit 'Use this skill whenever the user mentions...' clause with concrete trigger phrases. Matches the level-5 anchor exactly; not level 4 because the 'when' clause is fully explicit rather than merely serviceable.

5 / 5

Trigger Term Quality

Natural user phrasing is comprehensive and includes synonyms: 'backlog of pull requests', 'review and merge PRs', 'clear the PR queue', "'merge what's safe'", 'check that nothing breaks after merging', 'hands PR handling to an agent while they are away', plus mode terms 'review only' / 'dry run'. Covers both 'PRs' and 'pull requests' with quoted vernacular; no common variation is missing, so it beats the level-4 anchor.

5 / 5

Distinctiveness Conflict Risk

The merge/autonomous niche ('merge what's safe', PR queue backlog) is distinct, but trigger phrases like 'asks to review and merge PRs' and the supported 'review only' mode overlap with a plain PR/code-review skill, so a bare 'review this PR' request could route here incorrectly. Mostly distinct with minor overlap against closely related review skills — the level-4 anchor — rather than the minimal-conflict level-5 anchor.

4 / 5

Total

19

/

20

Passed

Validation

100%

Checks the skill against the spec for correct structure and formatting. All validation checks must pass before discovery and implementation can be scored.

Validation — 16 / 16 Passed

Validation for skill structure

No warnings or errors.

Repository
administrakt0r/pro-skills-repo-administraktor
Reviewed

Table of Contents

Is this your skill?

If you maintain this skill, you can claim it as your own. Once claimed, you can manage eval scenarios, bundle related skills, attach documentation or rules, and ensure cross-agent compatibility.