CtrlK
BlogDocsLog inGet started
Tessl Logo

ax-java-agent

Use when writing Java code with `dev.axllm:ax` for agents, child delegation, tools, MCP, citations, persistent playbook learning, stage instructions, runtime state, final typed responses, and direct-respond executor skipping.

56

Quality

70%

Does it follow best practices?

Run evals on this skill

Adds up to 20 points to the overall score

View guide
SecuritybySnyk

Passed

No findings from the security scan

Fix and improve this skill with Tessl

tessl review fix ./packages/java/skills/ax-java-agent/SKILL.md
SKILL.md
Quality
Evals
Security

Quality

Content

57%Weight 40%Scale 1-5

Reviews the quality of instructions and guidance provided to agents. Good implementation is clear, handles edge cases, and produces reliable results.

A dense, information-rich reference that respects the reader's intelligence (no basic-concept padding, every paragraph is package-specific) but reads like compressed release notes rather than an actionable skill: one executable example, no sequenced workflow for common tasks, and all behavioral detail inlined in SKILL.md with referenced detail files absent from the bundle.

Suggestions

Add one or two complete, executable Java examples per major feature (child agent registration, MCP client with authorizeToolCall, streamingForward with try-with-resources) so the guidance is copy-paste ready rather than descriptive.

Move the dense behavioral specification (Astra session work, streaming semantics, MCP cancellation rules) into reference files (e.g., references/session-semantics.md) and keep SKILL.md as a short overview with well-signaled links, ensuring referenced files actually ship in the bundle.

Add a short sequenced workflow section ('to build an agent: define signature → attach runtime → register children → forward/stream') with an explicit validation checkpoint (e.g., run a no-key example to verify the call shape before inventing new syntax).

DimensionReasoningScore

Conciseness

The body is dense, package-specific prose with essentially no padding and no explanation of concepts Claude already knows — e.g., 'Queued and applied are different states. HTTP applies updates at a response boundary.' It is not a 5 because of noticeable redundancy: the flat function namespace rules are stated across three overlapping bullets ('A flat function... is called `<namespace>.<name>`', "'own', the default since 25.0.0, is the above", "A flat function without its own namespace uses `utils`"), and the 'as in TypeScript' qualifier is repeated to the point of noise.

4 / 5

Actionability

There is one complete executable snippet (the 'Core Pattern' `Ax.agent(...)`/`forward(...)` example) and scattered concrete API names (`AxAgent.addChildAgent`, `Ax.fn(name).namespace("crm")`, `actorMode: 'completion'`, `agent.streamingForward(...)`), but the bulk of the body is behavioral description ('A provisional answer is not successful completion while started tools remain unresolved', 'Cancellation propagates through a delegated child') rather than executable how-to guidance, and the referenced sources of exact syntax (`examples/`, `API.md`) are not present in the skill bundle. This sits between 'some concrete guidance but incomplete' (3) and 'mostly executable guidance' (4), and the missing referenced files and descriptive bias pull it to 3.

3 / 5

Workflow Clarity

The body is organized as topical sections (When To Use, Package Facts, Core Pattern, ... Guardrails) rather than a sequenced workflow — there is no step ordering for common tasks (e.g., 'create agent → register children → attach runtime → stream'), and checkpoints exist only implicitly in Guardrails ('Start from package examples for exact native syntax before inventing a new call shape', 'Treat AxIR as the source of generated package truth'). This matches the anchor 'steps listed but validation gaps; sequence present but checkpoints missing or implicit'; it is not a 2 because the sections and guardrails do give a coherent reading and error-avoidance structure.

3 / 5

Progressive Disclosure

Section headers are clear and the body stays at overview altitude in places (Relevant API Surface, Package Facts), but ~100 lines of dense behavioral specification that would belong in a reference file are inlined in SKILL.md, and the files the body points to (`API.md`, `axir-api.json`, `axir-capabilities.json`, `examples/`, `src/examples/java/generation/`) do not exist in the skill bundle, so navigation to the detail is not verifiable. This matches 'some structure but could be better organized; references present but not clearly signaled; content that should be separate is inline'.

3 / 5

Total

13

/

20

Passed

Description

71%Weight 40%Scale 1-5

Based on the skill's description, can an agent find and select it at the right time? Clear, specific descriptions lead to better discovery.

A well-scoped, package-anchored description with an explicit 'Use when' trigger and a broad, mostly concrete capability list. Its weaknesses are the jargon-heavy trigger terms that users would rarely say naturally and the noun-list phrasing of capabilities instead of explicit actions.

Suggestions

Lead with one or two explicit action verbs (e.g., 'Create and configure Java RLM agents with dev.axllm:ax...') so the 'what' is stated as a concrete action rather than a feature noun list.

Trim internal jargon trigger terms (e.g., 'direct-respond executor skipping', 'stage instructions') or pair them with natural user phrasings (e.g., 'child agents', 'MCP clients', 'citations') to improve natural trigger coverage.

Add a few common synonyms or variations users might type (e.g., 'Ax Java SDK', 'axllm') to round out trigger-term coverage.

DimensionReasoningScore

Specificity

The description names a concrete domain ("writing Java code with `dev.axllm:ax`") and enumerates many specific capabilities ("agents, child delegation, tools, MCP, citations, persistent playbook learning, stage instructions, runtime state, final typed responses, and direct-respond executor skipping"). It falls short of a 5 because these are feature nouns rather than concrete action verbs, so coverage is broad but not phrased as actions; it is clearly above a 3, which expects only 1-2 concrete actions.

4 / 5

Completeness

The 'when' is explicit and specific ("Use when writing Java code with `dev.axllm:ax`"), and the 'what' is conveyed through the enumerated feature list. It is not a 5 because the 'what' is a noun list of feature areas rather than an explicit statement of what the skill does (e.g., 'Create and configure RLM agents...'), leaving the action partially implied.

4 / 5

Trigger Term Quality

Some genuinely natural trigger terms are present ("Java", "agents", "tools", "MCP", "citations", "runtime state"), but they are mixed with heavy internal jargon users would not naturally say ("persistent playbook learning", "child delegation", "stage instructions", "direct-respond executor skipping") and no synonyms or variations are offered. This matches the anchor 'Some relevant keywords but missing common variations or synonyms' rather than the level-4 'good keyword coverage with only a few natural terms missing'.

3 / 5

Distinctiveness Conflict Risk

The trigger is scoped to a specific named package (`dev.axllm:ax`) plus Java, giving it a clear niche with distinct triggers and minimal conflict risk with generic Java or other agent-framework skills. The specificity of the package identifier makes it unlikely to fire for the wrong skill.

5 / 5

Total

16

/

20

Passed

Validation

93%

Checks the skill against the spec for correct structure and formatting. All validation checks must pass before discovery and implementation can be scored.

Validation — 15 / 16 Passed

Validation for skill structure

CriteriaDescriptionResult

frontmatter_unknown_keys

Unknown frontmatter key(s) found; consider removing or moving to metadata

Warning

Total

15

/

16

Passed

Repository
ax-llm/ax
Reviewed

Table of Contents

Is this your skill?

If you maintain this skill, you can claim it as your own. Once claimed, you can manage eval scenarios, bundle related skills, attach documentation or rules, and ensure cross-agent compatibility.