CtrlK
BlogDocsLog inGet started
Tessl Logo

nc-hermes-e2e

Verifies Hermes skill discovery and fresh-session execution

60

Quality

75%

Does it follow best practices?

Run evals on this skill

Adds up to 20 points to the overall score

View guide
SecuritybySnyk

Passed

No findings from the security scan

Fix and improve this skill with Tessl

tessl review fix ./test/e2e/fixtures/hermes-skill-runtime/SKILL.md
SKILL.md
Quality
Evals
Security

Quality

Content

100%Weight 40%Scale 1-5

Reviews the quality of instructions and guidance provided to agents. Good implementation is clear, handles edge cases, and produces reliable results.

The body is a tight, single-purpose e2e test canary that gives unambiguous, fully actionable instructions with no padding and no need for external files. It exemplifies the simple-skill exception across conciseness, actionability, workflow clarity, and progressive disclosure.

DimensionReasoningScore

Conciseness

The body is lean and efficient — a heading plus a few short directives ('Reply with exactly PONG', env var, canary handling) with no padding or explanation of concepts Claude already knows, matching the score-5 anchor where every token earns its place; the SPDX comments are standard license boilerplate, not verbosity, so it stays at 5 rather than 4.

5 / 5

Actionability

The guidance is fully concrete and executable: 'do not use tools', 'Reply with exactly PONG and nothing else', and a specific env var, matching the score-5 anchor of fully executable, specific instruction; per the scoring notes, the absence of code in an instruction-only skill is not penalized when guidance is this actionable, so it does not drop to 4.

5 / 5

Workflow Clarity

This is a single-purpose skill under 50 lines whose single action (reply PONG, no tools) is unambiguous, so the simple-skill exception lets workflow clarity score 5; it is not a destructive or batch operation, so the validation cap does not apply, and it does not drop to 4 because there are no minor gaps in the one required action.

5 / 5

Progressive Disclosure

The skill is under 50 lines, needs no external references, and no references/scripts/assets bundle directories exist to verify, so per the simple-skill exception progressive disclosure scores 5 with a well-organized heading and clearly grouped directives; it does not drop to 4 because there is no organization gap to trim.

5 / 5

Total

20

/

20

Passed

Description

38%Weight 40%Scale 1-5

Based on the skill's description, can an agent find and select it at the right time? Clear, specific descriptions lead to better discovery.

The description states a clear, specific purpose but is written in internal jargon with no natural user trigger terms and no explicit 'Use when...' guidance, which caps completeness. It is distinct within its niche yet would not surface naturally from a user request.

Suggestions

Add an explicit 'Use when...' clause so the trigger condition is stated, not implied, which would raise completeness.

Replace or supplement the internal jargon ('fresh-session execution') with natural-language trigger phrases a user would actually say.

Add a second concrete action verb beyond 'Verifies' to broaden the specificity of stated capabilities.

DimensionReasoningScore

Specificity

The phrase 'Verifies Hermes skill discovery and fresh-session execution' names a clear domain (Hermes skill discovery / fresh-session execution) plus one concrete action ('Verifies'), matching the score-3 anchor of domain + 1-2 concrete actions without comprehensive coverage; it is not a 2 because the action and target are specific rather than minimal/generic, and not a 4 because only a single action verb is given rather than several.

3 / 5

Completeness

A clear 'what' is present ('Verifies Hermes skill discovery and fresh-session execution') but there is no 'Use when...' clause or equivalent trigger guidance, so per the capping rule completeness stays at 3; it is above 2 because the 'what' is explicit rather than vague, and below 4 because the 'when' is entirely absent.

3 / 5

Trigger Term Quality

The terms 'Hermes skill discovery' and 'fresh-session execution' are internal/test-tooling jargon a user would never naturally say, fitting the score-1 anchor of only technical jargon with no natural keywords; it does not rise to 2 because there is not even a generic natural phrase a user would utter.

1 / 5

Distinctiveness Conflict Risk

The niche is specific to Hermes e2e verification, giving it low overlap risk with unrelated skills, matching the score-4 anchor of mostly distinct with minor overlap risk; it does not reach 5 because it lacks distinct natural trigger phrases that would make wrong-skill triggering minimal.

4 / 5

Total

11

/

20

Passed

Validation

100%

Checks the skill against the spec for correct structure and formatting. All validation checks must pass before discovery and implementation can be scored.

Validation — 16 / 16 Passed

Validation for skill structure

No warnings or errors.

Repository
NVIDIA/NemoClaw
Reviewed

Table of Contents

Is this your skill?

If you maintain this skill, you can claim it as your own. Once claimed, you can manage eval scenarios, bundle related skills, attach documentation or rules, and ensure cross-agent compatibility.