CtrlK
BlogDocsLog inGet started
Tessl Logo

diplomatic-state-assessment

Use when assessing foreign state stability, predicting political upheaval, or advising allies on self-preservation. Follows Ji Zha's diplomatic mission model across Qi, Zheng, Wei, and Jin to evaluate leadership and identify safe havens.

61

Quality

71%

Does it follow best practices?

Run evals on this skill

Adds up to 20 points to the overall score

View guide

SecuritybySnyk

Passed

No findings from the security scan

Fix and improve this skill with Tessl

tessl review fix ./kg/ontology/ontology-v1/skus/procedural/skill_064/SKILL.md
SKILL.md
Quality
Evals
Security

Quality

Content

57%

Reviews the quality of instructions and guidance provided to agents. Good implementation is clear, handles edge cases, and produces reliable results.

A well-structured, single-purpose analytical skill with a clear step sequence, worked historical examples, and a validation checklist. Weaker on conciseness (redundant Overview/Expected Outcomes), actionability (high-level directives), and workflow_clarity (terminal rather than integrated validation).

Suggestions

Tighten redundancy: drop or shrink the Overview (it repeats the description) and fold 'Expected Outcomes' into the existing steps so every remaining token earns its place.

Make steps more concrete and actionable: replace abstract directives like 'Observe the ruling class behavior' with specific signals to look for (e.g., 'check whether power is consolidating into one family or office', 'note whether ministers retain land/forces').

Integrate validation into the workflow as a checkpoint after the Assess and Advise steps (validate → adjust advice) rather than leaving it only as a detached terminal checklist.

DimensionReasoningScore

Conciseness

The body is mostly efficient and avoids explaining concepts Claude already knows, but the Overview restates the frontmatter description and the 'Expected Outcomes' section ('Accurate prediction of political changes', 'Ability to advise allies on avoiding disaster') largely restates the steps, so it could be tightened. It is above the verbose score-1 anchor but not fully lean.

2 / 3

Actionability

Concrete directives and worked examples (Qi/Zheng/Wei/Jin cases) give genuine guidance, but individual steps like 'Observe the ruling class behavior' and 'Identify virtuous individuals in government' are high-level directives rather than concrete, specific procedures. This sits between the abstract score-1 anchor and the fully-copy-paste-ready score-3 anchor.

2 / 3

Workflow Clarity

A clear 4-step sequence (Assess → Evaluate → Advise → Identify safe havens) is present and a Validation section provides explicit verification checks, but validation is a detached terminal checklist rather than integrated feedback checkpoints, and interleaved non-sequential sections (Examples, Decision Points, Expected Outcomes) dilute the procedural flow. It exceeds the no-validation score-2 example but lacks the integrated feedback loops of the score-3 anchor.

2 / 3

Progressive Disclosure

The body is under 50 lines, requires no external references (none exist), and is organized into clearly labeled sections (Overview, Steps, Examples, Decision Points, Expected Outcomes, Validation), satisfying the rubric's simple-skill allowance for a score of 3 on well-organized sections alone.

3 / 3

Total

9

/

12

Passed

Description

85%

Based on the skill's description, can an agent find and select it at the right time? Clear, specific descriptions lead to better discovery.

A strong description: third-person voice, explicit 'Use when' triggers, a clear what/when split, and a highly distinctive niche. The main weakness is trigger-term breadth — the specialized historical framing is jargon users would not naturally invoke.

Suggestions

Add a few plain-language trigger variants users might actually say (e.g., 'analyzing a country's political risk', 'forecasting regime change', 'advising an ally on staying safe') alongside the current triggers to broaden natural-keyword coverage.

Move or soften the proper-noun jargon ('Ji Zha's diplomatic mission model across Qi, Zheng, Wei, and Jin') so the opening clause leads with the capability rather than the historical source.

DimensionReasoningScore

Specificity

The description lists multiple concrete domain actions — 'assessing foreign state stability, predicting political upheaval, or advising allies on self-preservation' plus 'evaluate leadership and identify safe havens' — matching the anchor for several specific actions rather than the single-action score-2 example.

3 / 3

Completeness

It explicitly answers both halves: an explicit 'Use when assessing... predicting... advising...' trigger (when) and 'Follows Ji Zha's diplomatic mission model... to evaluate leadership and identify safe havens' (what), satisfying the both-what-and-when anchor.

3 / 3

Trigger Term Quality

The 'Use when' clause surfaces reasonably natural phrases (foreign state stability, political upheaval, self-preservation), but the 'what' half leans on specialized jargon ('Ji Zha's diplomatic mission model across Qi, Zheng, Wei, and Jin') no user would naturally say, and coverage of common phrasing variations is moderate. It is above the no-keywords anchor (1) but below the broad-coverage anchor (3).

2 / 3

Distinctiveness Conflict Risk

The historical diplomatic-assessment niche (Ji Zha's missions to Qi, Zheng, Wei, Jin) is highly specific with distinct triggers and is very unlikely to overlap with other skills.

3 / 3

Total

11

/

12

Passed

Validation

100%

Checks the skill against the spec for correct structure and formatting. All validation checks must pass before discovery and implementation can be scored.

Validation16 / 16 Passed

Validation for skill structure

No warnings or errors.

Repository
baojie/shiji-kb
Reviewed

Table of Contents

Is this your skill?

If you maintain this skill, you can claim it as your own. Once claimed, you can manage eval scenarios, bundle related skills, attach documentation or rules, and ensure cross-agent compatibility.