CtrlK
BlogDocsLog inGet started
Tessl Logo

diagnose-agent-context-failures

Explain or summarize Baruch Sadogursky's JavaZone 2026 talk "The Right 300 Tokens Beat 100k Noisy Ones" and answer questions about its four context antipatterns, demos, conclusions, and evaluation caveats. Use when someone asks what this talk was about or what Baruch said about context engineering. Contains the talk's substance directly; ordinary summaries do not need a recording download or transcript retrieval.

76

Quality

95%

Does it follow best practices?

Run evals on this skill

Adds up to 20 points to the overall score

View guide
SecuritybySnyk

Passed

No findings from the security scan

SKILL.md
Quality
Evals
Security

Quality

Content

88%Weight 40%Scale 1-5

Reviews the quality of instructions and guidance provided to agents. Good implementation is clear, handles edge cases, and produces reliable results.

A well-engineered knowledge-brief skill: the process steps give concrete, executable guidance for matching and answering questions, and the talk substance is novel, structured, and source-linked. The only weaknesses are minor — a meta-commentary section that could be trimmed and a large inline brief that could be split for progressive disclosure.

DimensionReasoningScore

Conciseness

The talk brief is dense but consists of novel information about a September 2026 talk that Claude cannot already know, so most tokens earn their place. Minor over-explanation exists in the 'How the argument works' meta-commentary about summarizing the rhetoric, which could be trimmed — matching the score-4 anchor rather than 5.

4 / 5

Actionability

As an instruction-only knowledge skill, the guidance is concrete and executable: 'Give a concise summary by default; expand the relevant arguments or demos when asked', 'Attribute claims to the speaker', 'Do not fetch the recording or transcript for information already supplied here', and the rule to consult the linked source for exact quotations and say what was verified. Per the scoring notes, the absence of code is not penalized.

5 / 5

Workflow Clarity

The two-step process is clearly sequenced ('Process steps in order. Do not skip ahead') with an explicit mismatch checkpoint in Step 1 ('If the request concerns another delivery, identify the mismatch and finish here') and a verification rule for quotations in Step 2. No destructive or batch operations exist, so no validation cap applies.

5 / 5

Progressive Disclosure

The body is well organized into clearly headed sections (identity, four antipatterns, conclusion, sources) with one-level-deep external source links and no nested references. At roughly 195 lines the full talk brief lives inline in SKILL.md; for a knowledge skill this is defensible, but it is a minor organization gap versus the score-5 clear-overview-with-split-content structure.

4 / 5

Total

18

/

20

Passed

Description

100%Weight 40%Scale 1-5

Based on the skill's description, can an agent find and select it at the right time? Clear, specific descriptions lead to better discovery.

A model description for a niche knowledge skill: explicit third-person what and when clauses, natural trigger terms including the speaker's name and talk title, and comprehensive enumeration of the covered content. The trailing note about not needing transcript retrieval adds useful scoping without padding.

DimensionReasoningScore

Specificity

The description lists multiple specific concrete actions ('Explain or summarize', 'answer questions about its four context antipatterns, demos, conclusions, and evaluation caveats') whose enumerated subject areas comprehensively cover the talk's scope. It matches the score-5 anchor rather than 4 because there are no minor gaps in coverage of what the skill does.

5 / 5

Completeness

Both questions are explicitly answered: the first sentence states what the skill does (explain, summarize, answer questions with enumerated topics), and 'Use when someone asks what this talk was about or what Baruch said about context engineering' gives a concrete when-clause with explicit trigger phrases.

5 / 5

Trigger Term Quality

Natural trigger phrases a user would actually say are comprehensively covered: the speaker's name ('Baruch Sadogursky', 'what Baruch said'), the event ('JavaZone 2026'), the full talk title, 'context engineering', and 'what this talk was about'. These include the synonyms users would most naturally use when asking about this talk.

5 / 5

Distinctiveness Conflict Risk

The description is anchored to one specific talk, speaker, and event, forming a clear niche with distinct triggers (talk title, speaker name, event name). It also scopes activation away from generic context-engineering implementation requests, so conflict risk with other skills is minimal.

5 / 5

Total

20

/

20

Passed

Validation

100%

Checks the skill against the spec for correct structure and formatting. All validation checks must pass before discovery and implementation can be scored.

Validation — 16 / 16 Passed

Validation for skill structure

No warnings or errors.

Repository
jbaruch/shownotes
Reviewed

Table of Contents

Is this your skill?

If you maintain this skill, you can claim it as your own. Once claimed, you can manage eval scenarios, bundle related skills, attach documentation or rules, and ensure cross-agent compatibility.