Content
75%Weight 40%Scale 1-5Reviews the quality of instructions and guidance provided to agents. Good implementation is clear, handles edge cases, and produces reliable results.
The content is a well-structured, highly actionable guide: concrete paths, config examples, commands, and a diagnostic table let Claude execute both analysis and profiling workflows immediately. All four dimensions land at 4 rather than 5 due to minor repetition, a few high-level suggestions, a missing error-recovery loop after validation, and content that could be offloaded to reference files.
Suggestions
Deduplicate the analysis guidance: fold the Mode 2 'Quick analysis shortcuts' into Mode 1's script-writing step (or reference it) so kernel-breakdown and merge guidance appear once.
Add a brief error-recovery branch to Mode 2 Step 5 (e.g. what to check and revert if the loss curve diverges from baseline after an optimization).
Move the profile-config field details and the NPU differences into a reference file (e.g. references/npu.md) to keep SKILL.md a leaner overview, since no bundle files currently exist.
| Dimension | Reasoning | Score |
|---|---|---|
Conciseness | The body is dense and infrastructure-specific (file paths, config fields, output naming) with no padding explaining concepts Claude already knows. It is not 5 because there is noticeable repetition — kernel-time breakdown appears in Mode 1 step 3 and again in Mode 2's 'Quick analysis shortcuts', and merge guidance is stated twice. | 4 / 5 |
Actionability | Provides a component table with exact file paths, a complete YAML profile config, runnable bash and Python snippets, and a bottleneck-to-solutions mapping table. It is not 5 because some guidance remains high-level ('Increase compute/comm overlap', 'check for memory leaks') and commands use placeholders like '<model>.yaml'. | 4 / 5 |
Workflow Clarity | Both modes are clearly sequenced with numbered steps, and Mode 2 includes an explicit validation step ('Re-profile with the same config to compare before/after', 'Verify training correctness is preserved (loss matches baseline)'). It is not 5 because there is no error-recovery feedback loop for when validation fails (e.g. what to do if loss diverges after an optimization). | 4 / 5 |
Progressive Disclosure | A single-file skill with no bundle files present; the body is well-organized into clear sections (components, Mode 1, Mode 2, NPU) with no nested references. It is not 5 because at ~150 lines some content — the full profile-config field reference and the NPU differences — would be more appropriately split into reference files to keep SKILL.md a leaner overview. | 4 / 5 |
Total | 16 / 20 Passed |