Content
65%Weight 40%Scale 1-5Reviews the quality of instructions and guidance provided to agents. Good implementation is clear, handles edge cases, and produces reliable results.
The body is dense, actionable, and well-organized with copy-paste-ready CLI commands and useful matrices. Weaknesses are the absence of validation checkpoints in the operational workflow and a monolithic structure with no progressive disclosure into bundle files.
Suggestions
Add validation checkpoints to the Common workflow — e.g. 'Confirm the callback is healthy (whoami/hostname return expected host) before lateral movement' and 'If no callback within N intervals, rebuild the payload and re-deliver.'
Split the agent/profile matrices and the Sliver/Cobalt Strike/Havoc comparison into reference files (e.g. AGENTS.md, PROFILES.md) referenced one level deep, keeping SKILL.md an overview.
Verify or correct the Python scripting example's imports against the actual mythic package so the snippet is copy-paste runnable.
| Dimension | Reasoning | Score |
|---|---|---|
Conciseness | Lean and action-oriented — bash commands, compact matrices, and a one-line intro with minimal padding; it assumes Claude knows C2 concepts. Not 5 because the 'Comparison vs Sliver / Cobalt Strike / Havoc' table and the intro 'Strengths:' sentence restate some context Claude largely already knows. | 4 / 5 |
Actionability | Mostly executable guidance — full install/build/profile CLI commands are copy-paste ready and cover the common cases, matching the 'mostly executable; minor gaps' anchor. Not 5 because the Python scripting snippet imports modules ('mythic_utilities', 'mythic_callbacks') whose names are questionable, making that part not reliably runnable. | 4 / 5 |
Workflow Clarity | A clear 9-step 'Common workflow' sequence is present (setup → redirector → build → deliver → survey → enum → lateral → persistence → cleanup), but it has no validation/verification checkpoints or feedback loops. Because this is a destructive C2 operational workflow, the rubric caps workflow clarity at 3. | 3 / 5 |
Progressive Disclosure | The skill is well-sectioned (Setup, Agent matrix, Profile matrix, Build, Tasking, OPSEC, Workflow, Comparison, References) but is a monolithic ~130-line body with no bundle files, and inline content (agent/profile matrices, comparison table) that could be split into reference files. The References section points to external URLs, not clearly-signaled local files. | 3 / 5 |
Total | 14 / 20 Passed |