CtrlK
BlogDocsLog inGet started
Tessl Logo

ax-rust-playbook

Use when writing Rust code with `axllm` for the playbook() context-engineering surface, agent-bound verified evolution, run-end learning, online updates, and rendering a playbook into a program.

63

Quality

79%

Does it follow best practices?

Run evals on this skill

Adds up to 20 points to the overall score

View guide
SecuritybySnyk

Passed

No findings from the security scan

Fix and improve this skill with Tessl

tessl review fix ./packages/rust/skills/ax-rust-playbook/SKILL.md
SKILL.md
Quality
Evals
Security

Quality

Content

72%Weight 40%Scale 1-5

Reviews the quality of instructions and guidance provided to agents. Good implementation is clear, handles edge cases, and produces reliable results.

A lean, well-structured orientation skill that respects token budget and delegates detail appropriately to package artifacts. Its weaknesses are an illustrative-only code pattern with undefined variables and the absence of an explicit, validated workflow for the evolve/rollback loop it advertises.

Suggestions

Make the Core Pattern self-contained by defining or sourcing 'llm', 'examples', and 'metric_fn' (e.g., reference a specific complete file under examples/), or show the minimal metric_fn signature so the pattern is copy-paste runnable.

Add a short numbered workflow for the agent-bound evolve loop (attach seed playbook -> run -> collect run-end failure signals -> evolve with verification -> rollback on failure), with an explicit verification checkpoint before persisting the evolved playbook.

Briefly state what 'mine grounded weaknesses' and 'verification' mean operationally (e.g., which method performs verification and what triggers rollback), since these terms currently carry the workflow's safety story without concrete backing.

DimensionReasoningScore

Conciseness

The 44-line body is lean and entirely package-specific: every line adds facts Claude would not know (Package Facts, the Core Pattern, the API surface, guardrails). There is no padding and no explanation of concepts Claude already knows, matching anchor 5 ('every token earns its place').

5 / 5

Actionability

The Core Pattern block is concrete Rust but references undefined variables ('llm', 'examples', 'metric_fn') and gives no detail on constructing the metric function or examples format, making it a template rather than executable code (anchor 3: 'missing key details'). The guardrail 'Start from package examples for exact native syntax' partially compensates by pointing to runnable examples, but the inline guidance itself is incomplete, below anchor 4's mostly-executable bar.

3 / 5

Workflow Clarity

A rough sequence is implicit in the When To Use bullets and the Core Pattern (create program, construct playbook, evolve), but the multi-step evolve loop — which the body itself says involves 'verification and exact rollback' — has no explicit step ordering or validation checkpoints. This matches anchor 3 ('sequence present but checkpoints missing or implicit'), not anchor 2 since the tasks are at least enumerated and anchored by code.

3 / 5

Progressive Disclosure

The skill is under 50 lines with no bundle files, and all detail is delegated one level deep to clearly named package artifacts ('API.md', 'axir-api.json', 'axir-capabilities.json', 'examples/') in a dedicated Package Facts section. Per the rubric's simple-skill guidance, this well-organized, cleanly split structure merits anchor 5.

5 / 5

Total

16

/

20

Passed

Description

78%Weight 40%Scale 1-5

Based on the skill's description, can an agent find and select it at the right time? Clear, specific descriptions lead to better discovery.

A strong description with an explicit trigger clause, a distinct package-specific niche, and good coverage of the skill's action surface. Its main weaknesses are jargon-dense phrasing ('context-engineering surface') and the absence of a standalone capability statement, leaving the 'what' implicit inside the 'Use when... for...' construction.

DimensionReasoningScore

Specificity

The description lists several specific actions ('playbook() context-engineering surface', 'agent-bound verified evolution', 'run-end learning', 'online updates', and 'rendering a playbook into a program'), giving broad coverage of the skill's capabilities. It sits between anchor 3 and 4: more actions than the 1-2 of anchor 3, but the jargon-heavy 'context-engineering surface' phrasing keeps it from anchor 5's fully concrete comprehensiveness.

4 / 5

Completeness

It has an explicit trigger ('Use when writing Rust code with `axllm`') and the 'what' is conveyed through the 'for...' enumeration of tasks, but there is no standalone 'what' statement (e.g., 'Provides...'). This matches anchor 4 — both what and when present, with the 'what' embedded and less explicit than anchor 5's clear two-part structure.

4 / 5

Trigger Term Quality

It includes natural terms a user of this package would say: 'Rust code', 'axllm', 'playbook', 'evolution', 'online updates'. A few natural terms are missing (e.g., 'ax', 'optimizer', 'GEPA'), so it matches anchor 4 ('good keyword coverage; a few natural terms missing') rather than anchor 5's synonym-and-extension-level coverage.

4 / 5

Distinctiveness Conflict Risk

The niche is unambiguous: Rust code with the `axllm` package around playbook construction, evolution, and rendering. The package-specific trigger terms make conflict with generic coding or document skills minimal, matching anchor 5 ('clear niche with distinct triggers; minimal conflict risk').

5 / 5

Total

17

/

20

Passed

Validation

93%

Checks the skill against the spec for correct structure and formatting. All validation checks must pass before discovery and implementation can be scored.

Validation — 15 / 16 Passed

Validation for skill structure

CriteriaDescriptionResult

frontmatter_unknown_keys

Unknown frontmatter key(s) found; consider removing or moving to metadata

Warning

Total

15

/

16

Passed

Repository
ax-llm/ax
Reviewed

Table of Contents

Is this your skill?

If you maintain this skill, you can claim it as your own. Once claimed, you can manage eval scenarios, bundle related skills, attach documentation or rules, and ensure cross-agent compatibility.