CtrlK
BlogDocsLog inGet started
Tessl Logo

ax-agent-rlm

This skill helps an LLM generate correct AxAgent RLM/runtime code using @ax-llm/ax. Use when the user asks about RLM code execution, AxJSRuntime, contextFields, contextPolicy, liveRuntimeState, promptLevel, stage prompt controls, executorModelPolicy, maxRuntimeChars, agent.test(...), llmQuery(...), recursionOptions, or long-running agent runtime behavior.

67

Quality

84%

Does it follow best practices?

Run evals on this skill

Adds up to 20 points to the overall score

View guide
SecuritybySnyk

Passed

No findings from the security scan

SKILL.md
Quality
Evals
Security

Quality

Content

85%Weight 40%Scale 1-5

Reviews the quality of instructions and guidance provided to agents. Good implementation is clear, handles edge cases, and produces reliable results.

The body is a highly actionable, framework-dense rulebook: executable code for every subsystem, exact defaults, explicit error-recovery feedback loops, and a validation harness. Its weaknesses are duplicated preset/model guidance repeated across three sections and a monolithic single-file layout that inlines several reference-grade sections instead of splitting them into bundled reference files.

Suggestions

Split reference-grade sections into bundled files (e.g., references/runtime-security.md, references/llmquery.md, references/test-harness.md) and keep SKILL.md as a tight overview with one-level-deep links, matching how the skill already links external examples.

Merge the preset guidance currently repeated in "Use These Defaults", "Context Policy Presets", and "Choosing Presets, Prompt Level, And Model Size" into a single decision table (task type -> preset/budget/model) to eliminate triplicated rules.

Fold the "Do Not Generate" section into the "RLM Actor Code Rules" section as negative examples next to their positive counterparts, since every entry there restates an earlier rule.

DimensionReasoningScore

Conciseness

The body is dense, imperative, framework-specific rules with no generic-concept padding (e.g., it never explains what an agent or runtime is), assuming Claude's competence throughout. It falls short of anchor 5 because preset guidance is repeated across "Use These Defaults", "Context Policy Presets", and "Choosing Presets, Prompt Level, And Model Size", and actor-code rules recur in "Do Not Generate" — trimming the duplication would tighten it.

4 / 5

Actionability

Fully executable guidance throughout: a copy-paste-ready canonical agent config, runtime security recipes (new AxJSRuntime({...}) with exact option shapes), a complete runnable agent.test(...) harness with imports, and concrete actor-turn JavaScript examples. Exact defaults are stated for every option (e.g., "maxSubAgentCalls ... Default is 100"), matching anchor 5.

5 / 5

Workflow Clarity

The mental model gives an explicit ordered pipeline (distiller -> executor -> responder), actor turns are a clear per-turn workflow with feedback loops ("Errors ... appear in Action Log; inspect them and fix the code on the next turn"; "If a result starts with [ERROR], inspect or branch on it"), and agent.test(...) provides an explicit validation checkpoint before full runs. The delegation decision guide acts as a checklist — matching anchor 5.

5 / 5

Progressive Disclosure

Sections are clearly headed and the skill links out to full working examples ("Fetch these for full working code"), but the 505-line body inlines content that clearly belongs in separate reference files — the runtime security options reference, the llmQuery API rules, the test-harness guide, and the option layout could each be a one-level-deep reference. With no bundle files at all, this matches anchor 3 (structure exists; content that should be separate is inline).

3 / 5

Total

17

/

20

Passed

Description

78%Weight 40%Scale 1-5

Based on the skill's description, can an agent find and select it at the right time? Clear, specific descriptions lead to better discovery.

The description answers both what and when explicitly, with a strong, API-exact trigger list that mirrors the terms a user of this framework would actually say. Its main weakness is that the 'what' is a single bundled action rather than several enumerated capabilities, and one trigger phrase is broad enough to slightly overlap sibling skills.

Suggestions

Enumerate 2-3 distinct capabilities in the 'what' clause (e.g., 'configure context policies and stage models, generate runtime-hardened actor code, and validate snippets with agent.test(...)') to raise specificity from one bundled action to several concrete actions.

Add one or two natural-language synonyms alongside the API identifiers (e.g., 'agent runtime loop' or 'runtime code sessions') so users who don't remember exact option names still match the trigger list.

Narrow 'long-running agent runtime behavior' to RLM-specific phrasing (e.g., 'long-running RLM actor sessions') to reduce overlap with the sibling ax-agent skill.

DimensionReasoningScore

Specificity

"helps an LLM generate correct AxAgent RLM/runtime code using @ax-llm/ax" names the domain plus one concrete action (generate code), but does not list several distinct actions — the remaining terms are triggers, not stated capabilities. This matches anchor 3 (domain + 1-2 concrete actions); anchor 4 would require multiple enumerated specific actions.

3 / 5

Completeness

"This skill helps an LLM generate correct AxAgent RLM/runtime code using @ax-llm/ax" clearly states what it does, and "Use when the user asks about RLM code execution, AxJSRuntime, ..." explicitly states when to use it with concrete trigger phrases — both halves answered explicitly, matching anchor 5.

5 / 5

Trigger Term Quality

"RLM code execution, AxJSRuntime, contextFields, contextPolicy, liveRuntimeState, promptLevel, stage prompt controls, executorModelPolicy, maxRuntimeChars, agent.test(...), llmQuery(...), recursionOptions" provides good, API-exact keyword coverage that users of this SDK would naturally type. It falls short of anchor 5 because the list is purely technical identifiers with no synonyms or broader natural-language variants.

4 / 5

Distinctiveness Conflict Risk

Niche API identifiers like "AxJSRuntime", "contextPolicy", and "agent.test(...)" create a clear niche with minimal conflict risk, but the broader trigger "long-running agent runtime behavior" and the close relationship to the sibling ax-agent skill family leave minor overlap risk — anchor 4 rather than 5.

4 / 5

Total

16

/

20

Passed

Validation

87%

Checks the skill against the spec for correct structure and formatting. All validation checks must pass before discovery and implementation can be scored.

Validation — 14 / 16 Passed

Validation for skill structure

CriteriaDescriptionResult

skill_md_line_count

SKILL.md is long (505 lines); consider splitting into references/ and linking

Warning

frontmatter_unknown_keys

Unknown frontmatter key(s) found; consider removing or moving to metadata

Warning

Total

14

/

16

Passed

Repository
ax-llm/ax
Reviewed

Table of Contents

Is this your skill?

If you maintain this skill, you can claim it as your own. Once claimed, you can manage eval scenarios, bundle related skills, attach documentation or rules, and ensure cross-agent compatibility.