CtrlK
BlogDocsLog inGet started
Tessl Logo

pr-review

Verify the technical accuracy of a vscode-docs pull request against the VS Code source code in microsoft/vscode and microsoft/vscode-copilot-chat. Use when reviewing a docs PR for factual correctness — setting names, command IDs, default values, API shapes, keybindings, versioned availability, and described behavior.

71

Quality

89%

Does it follow best practices?

Run evals on this skill

Adds up to 20 points to the overall score

View guide
SecuritybySnyk

Low

Low-risk findings worth noting

SKILL.md
Quality
Evals
Security

Quality

Content

88%Weight 40%Scale 1-5

Reviews the quality of instructions and guidance provided to agents. Good implementation is clear, handles edge cases, and produces reliable results.

A well-crafted procedure skill: it gives copy-paste-ready gh commands, per-category search strategies, deterministic categorization rules, and a complete report format, with clear error-recovery paths for unverifiable claims. The only weaknesses are mild — some repeated caveats that could be trimmed, and all content inline in one long file rather than splitting the report template into a reference.

Suggestions

De-duplicate the 'technical accuracy only — use release-note-writer/frontmatter-description skills' caveat (stated in both the intro and Step 6) and the gh-cli-powershell.md memory pointer (stated in both 'Repos to Check' and Step 3); state each once.

Consider moving the findings-report template and the claims-category table into a references/ file (e.g., references/report-format.md) to keep SKILL.md as a leaner overview, which would also improve progressive disclosure.

DimensionReasoningScore

Conciseness

The body is dense and assumes Claude's competence — no concept explanations (nothing teaches what a PR or a keybinding is), and guidance like 'Prefer one targeted lookup per claim — do not download full files when a search will do' is maximally compressed. Minor trimmable redundancy keeps it below 5: the 'technical accuracy only, use other skills' caveat appears both in the intro and again in the Summary section, the `gh-cli-powershell.md` memory pointer is repeated twice, and 'do not push commits or post review comments' overlaps with the Notes framing.

4 / 5

Actionability

Guidance is fully executable: exact commands with flags ('gh pr view <number> --json number,title,headRefName,baseRefName,files,body', 'gh search code --repo microsoft/vscode '"<exact-string>"'', 'gh api repos/microsoft/vscode/contents/<path>?ref=main'), concrete search starting points per claim category, a severity rubric with decision criteria, and a copy-paste report template. Specific examples cover the common cases (settings, commands, keybindings, API, chat tools, version availability) with an explicit fallback (gh api) for the known gh-search quoting pitfall.

5 / 5

Workflow Clarity

A clear six-step sequence (identify diff → extract claims → verify → categorize → report → summarize) with explicit validation/error-recovery checkpoints: cosmetic files are skipped up front, a claim that cannot be verified is routed to 'Unverified' rather than failed ('mark it Unverified rather than failing it — the author may have access to context the source does not expose'), verification failure has a defined handling path, and the verdict rule is deterministic ('any Errors → Needs changes'). This is a read-only review, so the destructive/batch cap does not apply.

5 / 5

Progressive Disclosure

No bundle files exist and every section is well-organized with clear headers, tables, and a fenced template, so navigation is easy and nothing is buried. It falls short of 5 because the skill is a single ~145-line monolith with no content split at all — the output report template (~30 lines) and the claims-category table could live in a references/ file if the token budget mattered — and the under-50-line simple-skill exception does not apply at this length.

4 / 5

Total

18

/

20

Passed

Description

87%Weight 40%Scale 1-5

Based on the skill's description, can an agent find and select it at the right time? Clear, specific descriptions lead to better discovery.

A strong description: it states a concrete, repo-specific verification task and pairs it with an explicit 'Use when...' clause enumerating exactly the kinds of factual claims it checks. Third person throughout, no fluff or over-claims; the only room for improvement is broader synonym coverage (e.g., 'fact-check', 'audit', 'drift') and naming enterprise policies among the checkable categories.

DimensionReasoningScore

Specificity

The description names the domain ('verify the technical accuracy of a vscode-docs pull request against the VS Code source code in microsoft/vscode and microsoft/vscode-copilot-chat') and enumerates concrete checkable claim types ('setting names, command IDs, default values, API shapes, keybindings, versioned availability, and described behavior'), giving several specific items with only minor coverage gaps (e.g., enterprise policies are checked by the skill but not named). It is not a clear 5 because it describes essentially one action (verification) rather than multiple distinct concrete actions like the 5-anchor's 'extract, fill forms, merge, convert' example.

4 / 5

Completeness

Both questions are answered explicitly and concretely: the 'what' is 'Verify the technical accuracy of a vscode-docs pull request against the VS Code source code in microsoft/vscode and microsoft/vscode-copilot-chat', and the 'when' is the explicit trigger clause 'Use when reviewing a docs PR for factual correctness — setting names, command IDs, default values, API shapes, keybindings, versioned availability, and described behavior.' This directly matches the 5-anchor pattern (action sentence followed by a concrete 'Use when...' clause); it is not a 4 because the 'when' clause is specific and enumerates concrete triggers rather than being merely adequate.

5 / 5

Trigger Term Quality

Natural trigger phrases are present — 'Use when reviewing a docs PR for factual correctness' plus explicit terms like 'verify', 'technical accuracy', and named repos — which map well to what a user would say ('fact-check this docs PR', 'review this PR against source'). Not a 5 because common user-side synonyms like 'fact-check', 'validate', or 'drift'/'audit' are not included, and the anchor-5 bar of comprehensive synonym coverage (cf. 'PDF files, PDFs, forms, .pdf') is not fully met.

4 / 5

Distinctiveness Conflict Risk

The niche is unambiguous: verifying vscode-docs PRs against two named source repos, with an explicit scope of 'technical accuracy' — it is clearly distinguishable from general code-review, style, or frontmatter skills and would not trigger for the wrong skill. It is not a 4 because the trigger is repo- and task-specific, so even overlap with closely related skills is minimal.

5 / 5

Total

18

/

20

Passed

Validation

93%

Checks the skill against the spec for correct structure and formatting. All validation checks must pass before discovery and implementation can be scored.

Validation — 15 / 16 Passed

Validation for skill structure

CriteriaDescriptionResult

frontmatter_unknown_keys

Unknown frontmatter key(s) found; consider removing or moving to metadata

Warning

Total

15

/

16

Passed

Repository
microsoft/vscode-docs
Reviewed

Table of Contents

Is this your skill?

If you maintain this skill, you can claim it as your own. Once claimed, you can manage eval scenarios, bundle related skills, attach documentation or rules, and ensure cross-agent compatibility.