CtrlK
BlogDocsLog inGet started
Tessl Logo

talk-martinelli-spec-driven-development

Answers questions about, summarises key insights from, and helps apply concepts from Simon Martinelli's talk "Lessons from Spec-driven Development" — providing verbatim-grounded explanations, audits, and artifact drafts. Use when the user asks about the AI Unified Process, system use cases as specs (vs user stories), self-contained systems vs microservices, skills/MCP servers/guardrails, AI-assisted ERP modernization, drift management, how architecture style impacts AI coding agents, or applying his spec-driven approach to current work.

74

Quality

93%

Does it follow best practices?

Run evals on this skill

Adds up to 20 points to the overall score

View guide
SecuritybySnyk

Passed

No findings from the security scan

SKILL.md
Quality
Evals
Security

Quality

Content

88%Weight 40%Scale 1-5

Reviews the quality of instructions and guidance provided to agents. Good implementation is clear, handles edge cases, and produces reliable results.

A well-structured instruction skill: every task workflow is sequenced from a common LOOKUP step, grounded with explicit citation and fallback rules, and ends in concrete deliverables. The main improvement opportunities are deduplicating the lookup instructions, moving the inlined garbled transcript quote into quote.md, and trimming the repeated 'safe excerpts' phrasing.

Suggestions

State the lookup procedure once: drop grounding rule 1 and keep only the LOOKUP block (or vice versa), and reference it from each task section instead of restating the quote/outline/transcript path.

Move the verbatim garbled quote in 'Draft an artifact' ("we have use case and hence you believe that for right...") into quote.md with a stable anchor, and cite it there rather than inlining transcript content in SKILL.md.

Consolidate the repeated 'quote short, non-sensitive excerpts' phrasing into the grounding rules once, letting task sections say 'safe excerpts' without re-explaining.

DimensionReasoningScore

Conciseness

The body is lean and assumes Claude's competence — no concept tutorials, just rules and procedures. However, the lookup procedure is stated twice (grounding rule 1: "read `outline.md` to locate the relevant section, then read that section of `transcript.md`" vs. the LOOKUP block repeating the same steps), and "quote short, non-sensitive excerpts" is repeated across several sections. These are minor trims, matching anchor 4 ('efficient; minor instances of over-explanation that could be trimmed') rather than anchor 5's 'every token earns its place'.

4 / 5

Actionability

Fully actionable for an instruction-only skill: numbered steps per task type, an enumerated audit checklist of eight dimensions, a concrete artifact template ("Actor, Preconditions, Main success scenario, Alternative flows, Postconditions"), explicit marking conventions ("[not from talk — added as a starting placeholder]", "[likely: X]"), and defined fallbacks ("If the answer genuinely isn't in the transcript, say so explicitly"). Per the rubric's scoring note, absence of code is not penalized when the guidance is this specific; it fits anchor 5's 'specific examples cover the common cases'.

5 / 5

Workflow Clarity

Every workflow shares the mandated LOOKUP first step, then clearly sequenced numbered steps with explicit validation checkpoints: the quote.md sufficiency decision ("If sufficient, use those. Otherwise open `outline.md`..."), grounding checks ("If a claim isn't in `transcript.md`, say 'the talk doesn't address this'"), and honest-exit paths ("If the framework genuinely doesn't fit the user's situation... say so"). This matches anchor 5: clear sequence, explicit validation, and feedback loops; no destructive/batch cap applies.

5 / 5

Progressive Disclosure

Good structure: a ~90-line overview pointing one level deep to three well-signaled files (`quote.md`, `outline.md`, `transcript.md`) with distinct roles, navigated by the LOOKUP procedure and the "Key quotes" section. It falls short of anchor 5 because the garbled verbatim quote inlined in "Draft an artifact" ("we have use case and hence you believe that for right...") duplicates transcript content that belongs in `quote.md`/`transcript.md`, and no bundle files were provided to verify the referenced paths resolve — anchor 4's 'minor organization gaps' fits better.

4 / 5

Total

18

/

20

Passed

Description

96%Weight 40%Scale 1-5

Based on the skill's description, can an agent find and select it at the right time? Clear, specific descriptions lead to better discovery.

A strong description: third-person voice, multiple concrete capabilities, and an explicit 'Use when...' clause with specific, natural trigger phrases. The only weakness is minor trigger overlap with broader spec-driven-development and MCP-server skills.

DimensionReasoningScore

Specificity

"Answers questions about, summarises key insights from, and helps apply concepts from", plus "providing verbatim-grounded explanations, audits, and artifact drafts" lists multiple concrete actions with comprehensive coverage of this skill's scope. It clearly exceeds anchor 4 ('several specific actions; minor gaps') — no capability of the skill body (Q&A, summarizing, applying, auditing, artifact drafting) is missing from the description.

5 / 5

Completeness

Both questions are explicitly answered: 'what' via "Answers questions about, summarises key insights from, and helps apply concepts from... providing verbatim-grounded explanations, audits, and artifact drafts" and 'when' via the concrete "Use when the user asks about..." clause listing trigger phrases. This matches anchor 5 exactly; the 'when' clause is explicit, not weakly implied (anchor 3/4 territory).

5 / 5

Trigger Term Quality

"AI Unified Process, system use cases as specs (vs user stories), self-contained systems vs microservices, skills/MCP servers/guardrails, AI-assisted ERP modernization, drift management" comprehensively covers the natural terms a user would say, including synonyms and contrast pairings. It is not merely 'good coverage with a few missing' (anchor 4); secondary topics in the body (team structure, reflection pipeline) are genuinely minor and reachable via the talk title and speaker name, which are both present.

5 / 5

Distinctiveness Conflict Risk

The talk-specific framing ("Simon Martinelli's talk \"Lessons from Spec-driven Development\"", "AI Unified Process") creates a clear niche with distinct triggers, but "skills/MCP servers/guardrails" and "drift management" could overlap with general MCP-server-development or spec-driven-development (e.g., spec-kit style) skills. This fits anchor 4 ('mostly distinct; minor overlap risk') better than anchor 5's 'minimal conflict risk'.

4 / 5

Total

19

/

20

Passed

Validation

93%

Checks the skill against the spec for correct structure and formatting. All validation checks must pass before discovery and implementation can be scored.

Validation — 15 / 16 Passed

Validation for skill structure

CriteriaDescriptionResult

metadata_version

'metadata.version' is missing

Warning

Total

15

/

16

Passed

Repository
jscraik/Agent-Skills
Reviewed

Table of Contents

Is this your skill?

If you maintain this skill, you can claim it as your own. Once claimed, you can manage eval scenarios, bundle related skills, attach documentation or rules, and ensure cross-agent compatibility.