CtrlK
BlogDocsLog inGet started
Tessl Logo

judge-long-distance-military-campaign

Use when evaluating the feasibility of a long-distance surprise military attack. Assesses distance risks, intelligence reliability, supply line vulnerabilities, and contingency planning. Based on Qin's failed attack on Zheng as a cautionary example.

66

Quality

79%

Does it follow best practices?

Run evals on this skill

Adds up to 20 points to the overall score

View guide

SecuritybySnyk

Passed

No findings from the security scan

Fix and improve this skill with Tessl

tessl review fix ./kg/ontology/ontology-v1/skus/procedural/skill_026/SKILL.md
SKILL.md
Quality
Evals
Security

Quality

Content

71%Weight 40%Scale 1-5

Reviews the quality of instructions and guidance provided to agents. Good implementation is clear, handles edge cases, and produces reliable results.

The content is a compact, well-organized decision framework with a clear step sequence and an explicit validation checklist. Its main weakness is actionability: the guidance stays at the level of abstract directives rather than concrete, verifiable procedures.

Suggestions

Tighten actionability by giving each step a concrete, checkable criterion (e.g., 'List every territory crossed and flag any without transit agreement') instead of directives like 'Consider how many territories must be crossed'.

Convert the Validation section into an explicit checklist with pass/fail conditions and a 'do not proceed if any check fails' gate to strengthen the feedback loop.

Trim restated theme lines (e.g., 'Recognize that long-distance surprise attacks rarely succeed') and the Expected Outcomes section to push conciseness from efficient to fully lean.

DimensionReasoningScore

Conciseness

The body is lean and well-sectioned with little padding, but a few bullets ('Recognize that long-distance surprise attacks rarely succeed', the Expected Outcomes section) restate the theme, so it sits just above midpoint rather than fully lean.

4 / 5

Actionability

Guidance is framed as a decision checklist ('Evaluate claims about target vulnerability', 'Calculate the distance to target') which is concrete in intent but lacks the specific, verifiable detail of executable instructions, matching the 'some concrete guidance but incomplete' anchor.

3 / 5

Workflow Clarity

Six numbered steps are clearly sequenced and a dedicated Validation section supplies explicit checkpoints ('Verify...', 'Confirm...', 'Check...'), though it lacks true validate-fix-retry feedback loops, keeping it below 5.

4 / 5

Progressive Disclosure

The skill is a single ~50-line file with no need for external references and is organized into clear sections (Overview, Steps, Warning Signs, Expected Outcomes, Validation), satisfying the simple-skill exception for a top score.

5 / 5

Total

16

/

20

Passed

Description

87%Weight 40%Scale 1-5

Based on the skill's description, can an agent find and select it at the right time? Clear, specific descriptions lead to better discovery.

The description is strong: it pairs an explicit 'Use when' trigger with a concrete list of what the skill assesses, anchored to a specific cautionary example. Trigger and completeness are excellent; only minor specificity and keyword-synonym gaps keep it from a perfect profile.

DimensionReasoningScore

Specificity

Names the domain and lists several concrete assessment actions ('distance risks, intelligence reliability, supply line vulnerabilities, and contingency planning'), with only minor coverage gaps, fitting the 'several specific actions' anchor rather than the comprehensive 5.

4 / 5

Completeness

Explicitly answers both 'what' (Assesses distance risks, intelligence reliability, supply line vulnerabilities, contingency planning) and 'when' ('Use when evaluating the feasibility of a long-distance surprise military attack') with concrete trigger phrases.

5 / 5

Trigger Term Quality

Includes natural trigger phrasing ('Use when evaluating the feasibility of a long-distance surprise military attack') plus domain keywords a planner would say, though a few common synonyms are missing, placing it just below comprehensive.

4 / 5

Distinctiveness Conflict Risk

Occupies a clear niche (long-distance surprise military attack feasibility) with distinctive triggers tied to a specific historical example, minimizing overlap with other skills.

5 / 5

Total

18

/

20

Passed

Validation

100%

Checks the skill against the spec for correct structure and formatting. All validation checks must pass before discovery and implementation can be scored.

Validation16 / 16 Passed

Validation for skill structure

No warnings or errors.

Repository
baojie/shiji-kb
Reviewed

Table of Contents

Is this your skill?

If you maintain this skill, you can claim it as your own. Once claimed, you can manage eval scenarios, bundle related skills, attach documentation or rules, and ensure cross-agent compatibility.