Content
86%Weight 40%Scale 1-5Reviews the quality of instructions and guidance provided to agents. Good implementation is clear, handles edge cases, and produces reliable results.
A well-structured, highly actionable skill body: copy-paste commands with flags and outputs, clear per-command step sequences, front-loaded SSRF validation, a concrete error table, and an appropriately split one-level reference file. Weaknesses are minor: duplicated command listings, a hard-coded date in the title, and validation checkpoints that live in the error table rather than in the workflow steps themselves.
Suggestions
Remove the time-sensitive "(April 2026)" from the H1 and the repeated "Git for your SEO" tagline, and merge the Commands summary table with the per-command sections so each invocation appears only once — the current duplication is pure token cost.
Inline validation checkpoints into the step lists instead of relying on the Error Handling table: e.g., in compare, after fetching add a step "If fetch fails or returns an error status, report it and stop — do not run rules on partial data," mirroring the explicit SSRF validation step in baseline.
State what compare does when the current page returns a 4xx/5xx (the baseline workflow covers capturing error status, but compare behavior on an error-status current state is left implicit).
| Dimension | Reasoning | Score |
|---|---|---|
Conciseness | The body is almost entirely operational (commands, captured fields, rule severities, error actions) with no SEO-101 padding Claude already knows, but there is avoidable duplication: the Commands table repeats each invocation that the per-command Execution blocks give again, the "Git for your SEO" tagline appears twice (title and line 1), and the hard-coded date "(April 2026)" in the heading is time-sensitive token spend not placed in a deprecation section. This fits score 4 (efficient with minor instances that could be trimmed) better than 5, where every token would earn its place. | 4 / 5 |
Actionability | Every command is given as a copy-paste-ready invocation with real flags — "${CLAUDE_PLUGIN_ROOT}/scripts/claude-seo" run drift_baseline.py <url> --skip-cwv, drift_compare.py <url> --baseline-id 5, drift_history.py <url> --limit 10 — plus expected JSON outputs, a report-generation step, and a concrete per-scenario error table. This matches the score-5 anchor (fully executable, common cases covered); score 4 would require gaps in coverage or non-executable examples, which are absent. | 5 / 5 |
Workflow Clarity | Each command has a numbered step sequence with front-loaded validation ("Validate URL (SSRF protection via google_auth.validate_url())") and a thorough error-handling table with per-scenario actions. However, mid-sequence checkpoints are implicit rather than in the step list — e.g., the compare workflow doesn't state to halt if the fetch fails before running the 17 rules, and there is no explicit validate-retry loop — which fits score 4 (clear sequence, most checkpoints, minor validation gaps) rather than 5. | 4 / 5 |
Progressive Disclosure | The body is a well-sectioned overview (commands, captured elements, severity summary, storage, workflows) and correctly defers the full rule set to a clearly signaled, one-level-deep reference: "Load references/comparison-rules.md for the full rule set with thresholds, recommended actions, and cross-skill references" — and that file exists in the bundle with no further nesting. This matches the score-5 anchor (clear overview, well-signaled one-level-deep references, easy navigation). | 5 / 5 |
Total | 18 / 20 Passed |