CtrlK
BlogDocsLog inGet started
Tessl Logo

grill-me

Interview the user relentlessly about a plan or design until reaching shared understanding, resolving each branch of the decision tree. Use when user wants to stress-test a plan, get grilled on their design, or mentions "grill me".

68

Quality

82%

Does it follow best practices?

Run evals on this skill

Adds up to 20 points to the overall score

View guide

SecuritybySnyk

Passed

No findings from the security scan

SKILL.md
Quality
Evals
Security

Quality

Content

86%Weight 40%Scale 1-5

Reviews the quality of instructions and guidance provided to agents. Good implementation is clear, handles edge cases, and produces reliable results.

An exemplary lean, instruction-only skill: it assumes Claude's competence, gives concrete behavioral rules (recommended answer per question, explore the codebase when possible), and needs no external files. The only refinement worth making is making the terminal checkpoint explicit — restating the shared understanding or decision-tree resolution before finishing.

DimensionReasoningScore

Conciseness

The body is four sentences with zero padding — no explanation of concepts Claude already knows, no filler. Every token earns its place ("Interview me relentlessly... Walk down each branch... For each question, provide your recommended answer"), matching the lean-and-efficient anchor.

5 / 5

Actionability

For an instruction-only skill, the guidance is concrete: per-question behavior is specified ("provide your recommended answer"), sequencing is given ("resolving dependencies between decisions one-by-one"), and there is an explicit rule for when to explore the codebase instead of asking. It stops short of fully-executable guidance — e.g., no instruction to summarize or restate the shared understanding at the end — leaving minor gaps versus the level-5 anchor.

4 / 5

Workflow Clarity

The single interview task is unambiguous with an explicit stop condition ("until we reach a shared understanding") and per-question guidance, which under the simple-skill note could support a 5. However, the checkpoint is implicit — there is no explicit step to confirm or restate the shared understanding with the user before concluding — placing it at the clear-sequence-with-minor-gaps anchor.

4 / 5

Progressive Disclosure

The skill is under 50 lines, has no bundle files (references/, scripts/, assets/ are absent), and contains nothing that belongs in separate files — all guidance fits naturally in SKILL.md. Per the rubric's simple-skill guidance, this earns full credit; no references are buried because none are needed.

5 / 5

Total

18

/

20

Passed

Description

78%Weight 40%Scale 1-5

Based on the skill's description, can an agent find and select it at the right time? Clear, specific descriptions lead to better discovery.

A strong description: it uses third-person voice, states what the skill does, and includes an explicit 'Use when' clause with natural, distinctive trigger phrases including the memorable "grill me". The main gap is specificity — it describes only two concrete actions rather than enumerating the interview behaviors (recommended answers, dependency resolution) the body actually contains.

Suggestions

Enumerate 1-2 more concrete actions in the description, e.g., "proposes a recommended answer for each question and resolves dependencies between decisions" — this would lift specificity from the 1-2-actions anchor to several-actions.

Add one or two common natural synonyms to the trigger clause, such as "poke holes in my plan" or "play devil's advocate", to broaden trigger term coverage.

DimensionReasoningScore

Specificity

The description names the domain (plan/design stress-testing) and two concrete actions — "Interview the user relentlessly" and "resolving each branch of the decision tree" — matching the 1-2 concrete actions anchor. It does not list several specific actions (e.g., probing assumptions, recommending answers, resolving dependencies), so it falls short of the level-4 anchor.

3 / 5

Completeness

It explicitly answers both questions: what ("Interview the user relentlessly about a plan or design until reaching shared understanding, resolving each branch of the decision tree") and when, with concrete trigger phrases ("stress-test a plan", "get grilled on their design", "grill me"). This matches the level-5 anchor; it is not level 4 because the 'when' clause is already fully explicit rather than needing more specificity.

5 / 5

Trigger Term Quality

"Use when user wants to stress-test a plan, get grilled on their design, or mentions 'grill me'" provides good natural trigger phrases users would actually say. A few common variations are missing (e.g., "poke holes in my plan", "devil's advocate", "pressure-test", "challenge my design"), keeping it below the comprehensive-synonyms anchor.

4 / 5

Distinctiveness Conflict Risk

The distinctive "grill me" trigger and interview framing carve out a clear niche, but "stress-test a plan" has minor overlap risk with closely related plan-review or feedback skills. It is above level 3 (the trigger is distinctive, not merely 'somewhat specific') but below level 5 because plan/design review is a crowded space.

4 / 5

Total

16

/

20

Passed

Validation

100%

Checks the skill against the spec for correct structure and formatting. All validation checks must pass before discovery and implementation can be scored.

Validation — 16 / 16 Passed

Validation for skill structure

No warnings or errors.

Repository
openstatusHQ/data-table-filters
Reviewed

Table of Contents

Is this your skill?

If you maintain this skill, you can claim it as your own. Once claimed, you can manage eval scenarios, bundle related skills, attach documentation or rules, and ensure cross-agent compatibility.