CtrlK
BlogDocsLog inGet started
Tessl Logo

cross-examine

Interview the user relentlessly about a plan or design until reaching shared understanding, resolving each branch of the decision tree. Use when user wants to stress-test a plan, get grilled on their design, or mentions "grill me".

72

Quality

88%

Does it follow best practices?

Run evals on this skill

Adds up to 20 points to the overall score

View guide

SecuritybySnyk

Passed

No findings from the security scan

The canonical home for this skill is grill-me in mattpocock/skills

SKILL.md
Quality
Evals
Security

Quality

Content

92%Weight 40%Scale 1-5

Reviews the quality of instructions and guidance provided to agents. Good implementation is clear, handles edge cases, and produces reliable results.

An exemplary lean instruction-only skill body: fully actionable behavioral directives with a clear termination condition and no padding. The only improvement space is defining the end-of-interview output and optionally one example question.

DimensionReasoningScore

Conciseness

Four lean sentences, every one a directive, with zero padding and no explanations of concepts Claude already knows. The body adds behavior the description alone doesn't fully specify (one-at-a-time, recommendation per question, codebase-first), so every token earns its place — anchor 5, not 4, since nothing could be trimmed.

5 / 5

Actionability

'Ask the questions one at a time', 'For each question, provide your recommended answer', and 'If a question can be answered by exploring the codebase, explore the codebase instead' are concrete, unambiguous directives. Minor gaps keep it below anchor 5: no example question illustrating the desired depth, and no stated end-of-interview deliverable (e.g., summarize the resolved decisions).

4 / 5

Workflow Clarity

A single-purpose skill under 50 lines with an unambiguous loop — one question at a time, recommendation each round, resolve each branch — and an explicit termination condition ('until we reach a shared understanding'). No destructive or batch operations, so no validation cap applies; the simple-skill exception makes this a clear anchor 5.

5 / 5

Progressive Disclosure

Under 50 lines with no need for external references — no bundle files exist and none are referenced, correctly. At four sentences the content is trivially well-organized; per the under-50-lines guideline this is anchor 5, not 4, since there is nothing that should be split out.

5 / 5

Total

19

/

20

Passed

Description

83%Weight 40%Scale 1-5

Based on the skill's description, can an agent find and select it at the right time? Clear, specific descriptions lead to better discovery.

A concise, vivid, third-person description that clearly states both capability and explicit triggers. Its main weaknesses are incomplete synonym coverage and some overlap risk with general plan-review skills.

DimensionReasoningScore

Specificity

Lists several concrete actions — 'Interview the user relentlessly', 'resolving each branch of the decision tree', 'until reaching shared understanding' — with only minor coverage gaps (the recommend-an-answer-per-question behavior is absent). It goes beyond the 1–2 bare actions of anchor 3 but not to the comprehensive coverage of anchor 5.

4 / 5

Completeness

Explicitly answers both what ('Interview the user relentlessly about a plan or design... resolving each branch of the decision tree') and when with concrete trigger phrases ('Use when user wants to stress-test a plan, get grilled on their design, or mentions "grill me"'), matching the anchor 5 example structure exactly. The when-clause is already explicit and specific, so anchor 4's 'could be more explicit' does not apply.

5 / 5

Trigger Term Quality

'stress-test a plan', 'get grilled on their design', and the literal 'grill me' are phrases a user would naturally say. A few natural synonyms are missing ('poke holes in my plan', 'pressure-test', 'interview me about my design'), so it falls short of anchor 5's comprehensive synonym coverage.

4 / 5

Distinctiveness Conflict Risk

'grill me' / 'get grilled' are distinctive triggers creating a clear niche, but 'stress-test a plan' has real overlap risk with general plan-review and architecture-critique skills. Mostly distinct with minor overlap risk — anchor 4; not 5 because a plausible wrong-skill trigger exists.

4 / 5

Total

17

/

20

Passed

Validation

100%

Checks the skill against the spec for correct structure and formatting. All validation checks must pass before discovery and implementation can be scored.

Validation — 16 / 16 Passed

Validation for skill structure

No warnings or errors.

Repository
zebbern/claude-code-guide
Reviewed

Table of Contents

Is this your skill?

If you maintain this skill, you can claim it as your own. Once claimed, you can manage eval scenarios, bundle related skills, attach documentation or rules, and ensure cross-agent compatibility.