CtrlK
BlogDocsLog inGet started
Tessl Logo

ax-playbook

This skill helps an LLM generate correct playbook code using @ax-llm/ax. Use when the user asks about playbook(), AxPlaybook, context playbooks, evolving context, ACE / Agentic Context Engineering, agent.playbook(), or growing/applying task knowledge offline and online with evolve() and update().

68

Quality

86%

Does it follow best practices?

Run evals on this skill

Adds up to 20 points to the overall score

View guide
SecuritybySnyk

Passed

No findings from the security scan

SKILL.md
Quality
Evals
Security

Quality

Content

81%Weight 40%Scale 1-5

Reviews the quality of instructions and guidance provided to agents. Good implementation is clear, handles edge cases, and produces reliable results.

The body is highly actionable with executable patterns, clear sequences, validation gates, and a strong troubleshooting section. Its main weakness is repetition: the expensive-teacher gating and verify-gate semantics are restated across multiple sections, and multi-language port details inflate the single file.

Suggestions

Consolidate the isExpensive/teacherOptions gating (currently covered in 'Use These Defaults', the Agents teacher-options bullets, and Troubleshooting) into a single subsection, cross-referenced from the other spots.

Move the generated-packages table and the failureSignals/verify-gate paragraph for non-TypeScript ports into a separate reference file (e.g., references/generated-packages.md), keeping the TypeScript overview lean.

State the verify/rollback semantics once in the agent evolve bullet and trim the restated 'held-in improves, held-out within epsilon' wording from the generated-packages paragraph.

DimensionReasoningScore

Conciseness

Mostly efficient — the API facts (return shapes, lazy hydration, verify gate semantics) are non-obvious and earn their tokens. However, the isExpensive/teacherOptions gating is explained in at least four places ('Use These Defaults', the Agents teacher bullets, and Troubleshooting), and the verify-gate semantics are stated twice (Agents and the generated-packages paragraph). This repetition goes beyond 'minor instances that could be trimmed', matching 'Mostly efficient but includes some unnecessary explanation or could be tightened'.

3 / 5

Actionability

Fully executable copy-paste TypeScript examples cover the common cases: offline evolve with metric, online update with example/prediction/feedback, persist/restore via toJSON/load, and the agent playbook handle. The Troubleshooting section pairs concrete error messages with concrete fixes ('wrap them in example: { ... }'). This matches the top anchor.

5 / 5

Workflow Clarity

Multi-step workflows are clearly sequenced (create → evolve/update → applyTo → toJSON/load), the agent evolve verify gate provides an explicit validation checkpoint with rollback semantics, and the Troubleshooting section supplies error-recovery feedback loops for each listed failure mode. This matches 'Clear sequence with explicit validation steps; feedback loops for error recovery'.

5 / 5

Progressive Disclosure

No bundle files exist, so all content is inline in a single well-sectioned SKILL.md (~130 lines) with a clearly signaled cross-skill 'See Also' section. Structure is good and navigation is easy, but peripheral bulk — the five-language generated-packages table and the deep agent-evolve/failureSignals detail — is inlined where a reference file would keep the overview lean. This matches 'Good structure; most content is appropriately placed; minor organization gaps' rather than the well-split anchor 5.

4 / 5

Total

17

/

20

Passed

Description

87%Weight 40%Scale 1-5

Based on the skill's description, can an agent find and select it at the right time? Clear, specific descriptions lead to better discovery.

A strong description: it explicitly states what the skill does and provides a rich, natural set of trigger terms tied to the library's API names. Minor room to broaden coverage of related actions (persist/restore) and synonyms.

DimensionReasoningScore

Specificity

The description names the domain ('generate correct playbook code using @ax-llm/ax') and several concrete actions: 'growing/applying task knowledge offline and online with evolve() and update()'. It lists several specific operations, though coverage has minor gaps (no mention of persist/restore or applyTo), matching the anchor 'Lists several specific actions; minor gaps in coverage' rather than the comprehensive anchor 5.

4 / 5

Completeness

Both questions are explicitly answered: what — 'helps an LLM generate correct playbook code using @ax-llm/ax'; when — 'Use when the user asks about playbook(), AxPlaybook, ... evolve() and update()' with concrete trigger phrases. This matches the anchor 'Clearly and explicitly answers both what AND when with concrete trigger phrases'.

5 / 5

Trigger Term Quality

Trigger terms are strong and natural for this niche: 'playbook(), AxPlaybook, context playbooks, evolving context, ACE / Agentic Context Engineering, agent.playbook()'. A user of this library would naturally say these. It sits at 'Good keyword coverage; a few natural terms missing' (e.g., persist/snapshot/load phrasings) rather than the fully comprehensive anchor 5.

4 / 5

Distinctiveness Conflict Risk

The description carves out a clear niche (the @ax-llm/ax playbook/ACE feature) with distinct trigger terms, and it does not claim optimize()/GEPA territory that sibling skills cover. Minimal conflict risk with other skills.

5 / 5

Total

18

/

20

Passed

Validation

93%

Checks the skill against the spec for correct structure and formatting. All validation checks must pass before discovery and implementation can be scored.

Validation — 15 / 16 Passed

Validation for skill structure

CriteriaDescriptionResult

frontmatter_unknown_keys

Unknown frontmatter key(s) found; consider removing or moving to metadata

Warning

Total

15

/

16

Passed

Repository
ax-llm/ax
Reviewed

Table of Contents

Is this your skill?

If you maintain this skill, you can claim it as your own. Once claimed, you can manage eval scenarios, bundle related skills, attach documentation or rules, and ensure cross-agent compatibility.