CtrlK
BlogDocsLog inGet started
Tessl Logo

hermes-ephemeral-delegation

Trigger: broad exploration, multi-file reads, tests/builds, fresh review, or multi-step debug. Orchestrate complex work via delegate_task to protect context.

65

Quality

78%

Does it follow best practices?

Run evals on this skill

Adds up to 20 points to the overall score

View guide

SecuritybySnyk

Passed

No findings from the security scan

Fix and improve this skill with Tessl

tessl review fix ./internal/assets/skills/hermes-ephemeral-delegation/SKILL.md
SKILL.md
Quality
Evals
Security

Quality

Content

86%Weight 40%Scale 1-5

Reviews the quality of instructions and guidance provided to agents. Good implementation is clear, handles edge cases, and produces reliable results.

A tightly written orchestration skill: efficient token use, a decision table that makes routing unambiguous, and a well-placed, verified reference file. The two gaps are the absence of a worked delegate_task example and a recovery path when worker verification fails.

DimensionReasoningScore

Conciseness

The body is lean and directive: the Decision Gates table encodes seven situations in two columns, the mission checklist is a bare enumeration, and no section explains concepts Claude already knows. Every line is either a rule, a gate, a step, or a pointer. This matches the 'lean and efficient; every token earns its place' anchor rather than anchor 4, which reserves room for trimmable over-explanation — there is none here.

5 / 5

Actionability

The Decision Gates table and the six-item mission checklist give concrete, executable guidance (exact goal, file paths, constraints, expected evidence, allowed toolsets). It falls short of anchor 5 because there is no copy-paste-ready example of an actual delegate_task call or mission text — the reader must assemble one from the checklist.

4 / 5

Workflow Clarity

Execution Steps 1-6 are clearly sequenced with an explicit validation checkpoint (step 5, 'Verify the claimed output (check file existence, test result, side effect)') and the Output Contract covers discrepancies. It misses anchor 5 because there is no error-recovery feedback loop when verification fails — no instruction to retry, re-delegate, or fix — only reporting the discrepancy. The batch-parallel delegation context is covered by the 'batch parallel calls only for independent workstreams' rule, so the batch cap at 3 does not apply given validation is present.

4 / 5

Progressive Disclosure

The body is a concise overview with clear sections, and the single reference (references/tuning-knobs.md) exists, is one level deep, is clearly signaled with a description of what it contains ('Full table of delegate_task configuration parameters and the explicit toolset/MCP/skill checklist'). Content is appropriately split — the tuning table genuinely belongs in a reference file. This matches the 'clear overview with well-signaled one-level-deep references' anchor; the only reason to consider 4 would be wanting more references, but nothing in the body needs externalizing.

5 / 5

Total

18

/

20

Passed

Description

70%Weight 40%Scale 1-5

Based on the skill's description, can an agent find and select it at the right time? Clear, specific descriptions lead to better discovery.

A compact, trigger-first description with a genuine explicit trigger clause and a concrete mechanism (delegate_task). Its main weakness is that the 'what' side under-sells the skill's actual workflow (drafting self-contained missions, verifying worker self-reports, synthesizing results).

Suggestions

Expand the 'what' clause to name the full capability set, e.g. 'Orchestrate complex work via delegate_task: draft self-contained worker missions, verify worker self-reports, and synthesize results to protect the parent context.'

Add common trigger synonyms users would naturally say, such as 'PR review', 'refactoring across files', or 'codebase exploration', to broaden natural keyword coverage.

Sharpen distinctiveness by signaling the orchestration-vs-execution boundary in the description (e.g. 'for parent orchestrators, not delegated workers') to reduce overlap with plain review/build skills.

DimensionReasoningScore

Specificity

The description names the mechanism ('Orchestrate complex work via delegate_task') and a purpose ('protect context'), which is 1-2 concrete actions, but it does not enumerate the broader capability set (mission drafting, verification, synthesis) covered in the body. It sits above 'names the domain but actions are minimal' (2) because delegate_task and context protection are concrete, but below 4 because coverage of what the skill actually does is thin.

3 / 5

Completeness

Both halves are present: 'Trigger: ...' explicitly answers when, and 'Orchestrate complex work via delegate_task to protect context' answers what. The when clause is explicit with concrete trigger phrases, but the what is compressed — it omits the verify-and-synthesize loop that is central to the skill — so it falls short of the 'clearly and explicitly answers both' anchor 5 while clearly exceeding anchor 3.

4 / 5

Trigger Term Quality

Trigger phrases 'broad exploration, multi-file reads, tests/builds, fresh review, or multi-step debug' are natural terms users would say. Coverage is good but misses common synonyms like 'PR review', 'refactor', 'codebase mapping', or 'audit', keeping it below the comprehensive anchor 5.

4 / 5

Distinctiveness Conflict Risk

The delegate_task orchestration niche is distinctive and unlikely to be confused with file-format or analysis skills. Minor overlap risk remains because trigger terms like 'fresh review' and 'tests/builds' could also match dedicated review or build skills. It fits 'mostly distinct; minor overlap risk' rather than the minimal-conflict anchor 5.

4 / 5

Total

15

/

20

Passed

Validation

100%

Checks the skill against the spec for correct structure and formatting. All validation checks must pass before discovery and implementation can be scored.

Validation — 16 / 16 Passed

Validation for skill structure

No warnings or errors.

Repository
Gentleman-Programming/gentle-ai
Reviewed

Table of Contents

Is this your skill?

If you maintain this skill, you can claim it as your own. Once claimed, you can manage eval scenarios, bundle related skills, attach documentation or rules, and ensure cross-agent compatibility.