Content
71%Weight 40%Scale 1-5Reviews the quality of instructions and guidance provided to agents. Good implementation is clear, handles edge cases, and produces reliable results.
The body is an information-dense domain playbook with excellent actionability — concrete thresholds, worked dollar examples, and ready-to-use stakeholder language — and a well-signaled escalation/monitoring loop. Its weaknesses are length (it re-teaches concepts Claude knows and inlines ~90 lines of reference material that should live in separate files) and implicit rather than explicit mid-workflow validation checkpoints.
Suggestions
Split Core Knowledge (bill anatomy, market structures, PPA/REC mechanics) and the Edge Cases into references/ files (e.g. references/tariffs.md, references/ppa-evaluation.md), keeping SKILL.md as an overview with one-level-deep links.
Trim explanations of concepts Claude already knows (LMP components, REC definition, GHG Protocol Scope 2 basics) down to the project-specific numbers and caveats that matter.
Add explicit validation checkpoints inside 'How It Works' — e.g. after bid evaluation, 'reconcile the supplier's load model against your interval data before shortlisting', and after budget building, 'back-test the forecast against last year's actuals'.
| Dimension | Reasoning | Score |
|---|---|---|
Conciseness | Much of the ~220-line body is genuinely specialized (tariff ratchets, PLC tag mechanics, basis risk quantification, stacked payback math), but it also explains concepts Claude already knows — 'LMP = Energy Component + Congestion Component + Loss Component', '1 REC = 1 MWh of renewable generation attributes', GHG Protocol Scope 2 basics — and the whole could be tightened. This fits the score-3 anchor (mostly efficient, some unnecessary explanation, could be tightened) rather than score 2, since the majority of tokens carry domain data rather than padding. | 3 / 5 |
Actionability | For an instruction-only skill, guidance is fully executable: copy-paste-ready negotiation scripts ('ICE forward curves for 2027 are showing $42/MWh... your quote of $48/MWh reflects a 14% premium'), step-by-step ROI math ('Calculate current demand charges: Peak kW × demand rate × 12 months'), numeric decision thresholds, and a trigger/action/timeline escalation table. Specific worked examples (25-site PJM/ERCOT RFP, 500kW/2MWh battery payback, $35/MWh VPPA) cover the common cases, matching the score-5 anchor. | 5 / 5 |
Workflow Clarity | 'How It Works' gives a clear 6-step sequence, and the escalation table and performance-indicator red flags supply feedback checkpoints ('Wholesale prices exceed 2× budget assumption for 5+ consecutive days → ... Within 24 hours'). It falls short of score 5 because intermediate validation is implicit — e.g. verifying a bid model or budget forecast against actuals is referenced only via the after-the-fact KPI table, not as an explicit checkpoint inside the workflow. | 4 / 5 |
Progressive Disclosure | Section structure is good (Role, When to Use, How It Works, Core Knowledge, Decision Frameworks, Edge Cases, Communication, Escalation, KPIs), but everything is inlined in one ~30KB SKILL.md with no bundle files at all. The ~90 lines of Core Knowledge (bill anatomy, market structures, PPA types) clearly belong in reference files. This matches the score-3 anchor (some structure, content that should be separate is inline) rather than score 2, since headers make it navigable and no references are buried or nested. | 3 / 5 |
Total | 15 / 20 Passed |