CtrlK
BlogDocsLog inGet started
Tessl Logo

verify-rfc

Verify an RFC or design doc against the actual implementation — find drift, missing pieces, and undocumented changes. Trigger phrases include "verify the RFC", "compare design doc to implementation", "is the spec implemented", "does the code match the design", "what's drifted from the RFC", "audit the doc against", "spec compliance check", "is what we built what we designed".

68

Quality

86%

Does it follow best practices?

Run evals on this skill

Adds up to 20 points to the overall score

View guide
SecuritybySnyk

Low

Low-risk findings worth noting

SKILL.md
Quality
Evals
Security

Quality

Content

70%Weight 40%Scale 1-5

Reviews the quality of instructions and guidance provided to agents. Good implementation is clear, handles edge cases, and produces reliable results.

A well-sequenced, genuinely skeptical verification workflow with strong checklists and honest failure handling, but it is monolithic: two full report templates are inlined that should be split into reference files, and the admonition repetition adds tokens without adding guidance. Splitting the report templates out would raise both conciseness and progressive disclosure.

Suggestions

Move the two ~60-line report templates (main report and inconclusive report) into references/report-templates.md and reference them by link, which would address both the conciseness and progressive_disclosure gaps.

Consolidate the repeated skeptic admonitions ("BE SKEPTICAL", "BE CAREFUL", "CRITICAL") into a single statement of the skepticism principle in Core Principles, removing the duplicated emphasis markers.

Make the Phase 2 local-search guidance concrete by naming example search patterns (e.g., file-name globs and content greps for a requirement) instead of the generic description of grep/glob/read.

DimensionReasoningScore

Conciseness

Mostly efficient but the two full report templates (~60 lines each) are inlined wholesale and skeptical admonitions are repeated ("CRITICAL", "BE SKEPTICAL", "BE CAREFUL") where one statement of the principle would do. Not a 2 because there is no conceptual padding explaining things Claude already knows.

3 / 5

Actionability

Concrete tool invocations (read_document, code_search with quoted query patterns) plus explicit pass/fail rubrics (Match Quality, Evidence Strength, Fidelity) and enumerated false-positive rules make this mostly executable. A 5 would require fully copy-paste-ready commands; the local-search guidance ("use grep/glob/read") stays descriptive rather than specific.

4 / 5

Workflow Clarity

A clear 4-phase sequence (fetch → search → vet → report) with explicit validation checklists in Phase 3, an honest-inconclusive branch, and a troubleshooting section providing error-recovery feedback loops. This is not a destructive or batch operation, so no cap applies.

5 / 5

Progressive Disclosure

No bundle files exist, so everything lives in one ~215-line SKILL.md; the two large report templates clearly belong in a references/ file. Section structure itself is clear (not a 2), but content that should be separate is inline and no references are used at all (not a 4).

3 / 5

Total

15

/

20

Passed

Description

96%Weight 40%Scale 1-5

Based on the skill's description, can an agent find and select it at the right time? Clear, specific descriptions lead to better discovery.

An excellent description: concrete actions, explicit what-and-when structure, third-person voice, and an unusually thorough set of natural trigger phrases. The only weakness is minor: one incomplete trigger fragment and slight overlap territory with generic code-review triggers.

DimensionReasoningScore

Specificity

"Verify an RFC or design doc against the actual implementation — find drift, missing pieces, and undocumented changes" names multiple specific concrete actions with comprehensive coverage of the verification domain. Anchor 4 requires 'minor gaps in coverage' and no gaps are present.

5 / 5

Completeness

Both what (verify implementation against spec, find drift/missing pieces/undocumented changes) and when (explicit trigger-phrase clause) are clearly and explicitly stated, matching the top anchor exactly.

5 / 5

Trigger Term Quality

Eight natural user phrasings including synonyms and variations ("verify the RFC", "spec compliance check", "does the code match the design", "is what we built what we designed") comprehensively cover how a user would naturally request this.

5 / 5

Distinctiveness Conflict Risk

The spec-vs-implementation niche is clear and mostly distinct from code-review skills, but "audit the doc against" is a truncated fragment and a few phrases ("does the code match the design") have minor overlap risk with general review skills.

4 / 5

Total

19

/

20

Passed

Validation

100%

Checks the skill against the spec for correct structure and formatting. All validation checks must pass before discovery and implementation can be scored.

Validation — 16 / 16 Passed

Validation for skill structure

No warnings or errors.

Repository
gleanwork/claude-plugins
Reviewed

Table of Contents

Is this your skill?

If you maintain this skill, you can claim it as your own. Once claimed, you can manage eval scenarios, bundle related skills, attach documentation or rules, and ensure cross-agent compatibility.