Content
75%Weight 40%Scale 1-5Reviews the quality of instructions and guidance provided to agents. Good implementation is clear, handles edge cases, and produces reliable results.
An exemplary codebase-patterns skill body: dense with non-derivable project knowledge, cleanly sectioned, with a clearly sequenced pipeline flow, explicit critical rules, and honest boundary statements ('Treat streaming as an open gap', 'no hit-rate target is computed or enforced'). Weaknesses are minor — the cache section could be trimmed, and verification guidance lacks an explicit test command and feedback loop.
| Dimension | Reasoning | Score |
|---|---|---|
Conciseness | Nearly every line is codebase-specific fact Claude cannot know (paths, function names, feature flags, cache key structure), with no filler and no explanation of known concepts — 'Bindings are separate crates, not features of crates/xberg' is typical of the lean, knowledge-dense style. It stops short of the 5 'every token earns its place' anchor because the cache bullet carries multi-clause reasoning ('the commit SHA plus a hash of any uncommitted diff (build.rs), so it IS a build fingerprint') that could be trimmed. | 4 / 5 |
Actionability | Guidance is concrete and directly executable against the codebase: exact paths ('crates/xberg/src/core/pipeline/'), a runnable command ('scripts/sync_supported_counts.py verify'), precise rules ('Register above 50 to override a built-in'), and named entry points ('core::pipeline::run_pipeline(doc, config)'). Per the rubric's instruction-skill note, absent code snippets are not penalized when guidance is actionable; it is not a 5 because there are no copy-paste examples for the common cases (e.g., adding or modifying an extractor). | 4 / 5 |
Workflow Clarity | The 'Flow' section gives a clearly sequenced four-step pipeline (Detect → Route → Extract → Post-process), and the Verification section mandates testing 'both success and failure paths' plus keeping the count test in sync with 'FORMATS'. It matches the 4 anchor 'clear sequence with most checkpoints present; minor validation gaps' — the validation guidance names what to check but no explicit test command or feedback loop (run → fail → fix → re-run) for arbitrary pipeline changes. | 4 / 5 |
Progressive Disclosure | No bundle files exist (no references/, scripts/, or assets/), so the skill is judged on its own structure: eleven well-labeled sections with an appropriate overview/index role, and cross-navigation is one level deep via named sibling skills ('mime-detection-routing', 'plugin-architecture-patterns', 'feature-flag-policy', 'benchmark-workflow', 'test-corpus'). It matches the 4 anchor 'good structure; most content appropriately placed; minor organization gaps' — the extractor-module catalogue and features table are borderline content that could live in reference files if the skill grew. | 4 / 5 |
Total | 16 / 20 Passed |