CtrlK
BlogDocsLog inGet started
Tessl Logo

verify-worldmonitor

Verify WorldMonitor dashboard behavior with existing browser tests or a scoped manual drive. Use for panels, map layers, settings, search, country briefs, boot, or dashboard screenshots.

70

Quality

88%

Does it follow best practices?

Run evals on this skill

Adds up to 20 points to the overall score

View guide
SecuritybySnyk

Passed

No findings from the security scan

SKILL.md
Quality
Evals
Security

Quality

Content

85%Weight 40%Scale 1-5

Reviews the quality of instructions and guidance provided to agents. Good implementation is clear, handles edge cases, and produces reliable results.

A highly actionable, well-sequenced operational skill body with strong code, commands, validation, and feedback loops. Its main weakness is progressive disclosure: the body leans heavily on referenced feature/step files that are not present in the provided bundle.

Suggestions

Ship the referenced bundle files (features/README.md and the per-feature recipes, plus the steps/ directory including _profile.mjs) alongside SKILL.md so the clearly-signaled links actually resolve.

Inline a one-line summary of each feature recipe next to its link in the step table, so a reader can still scope the work when the features/ files are unavailable.

Trim the Launch isolation paragraph to the essential rule (one instance per worktree; never adopt a server this run did not start) to push conciseness from 4 toward 5.

DimensionReasoningScore

Conciseness

The body is dense and operational, explaining only app-specific non-obvious details (VITE_E2E markers, readiness probe, isolation rules) rather than concepts Claude already knows, with only minor spots that could be trimmed such as the detailed isolation paragraph. Not a 5 because a few explanatory passages are borderline over-elaborated.

4 / 5

Actionability

Provides copy-paste-ready code (a complete default-exported async step with the real harness API) and exact commands (launch/doctor/drive/cleanup with port and env flags), plus a step-to-feature table and a stable-handles list, covering the common cases fully.

5 / 5

Workflow Clarity

A clear sequence (Choose the proof -> Launch -> Doctor -> Drive -> Evidence -> Cleanup) with explicit validation (doctor before driving, exit code 0 = OK), feedback loops (failed drive -> doctor -> cleanup + relaunch), and a proof-standards checklist; validation is present so the destructive-operation cap does not apply.

5 / 5

Progressive Disclosure

Sections are well-organized and references are clearly signaled one level deep, but checked against the actual bundle most referenced paths (features/README.md, features/*.md, steps/_profile.mjs, steps/*.mjs) do not exist -- only scripts/drive.mjs and scripts/wm-verify.sh are present -- so the promised navigation to detail files is broken, a gap larger than the 'minor organization gaps' allowed at 4.

3 / 5

Total

17

/

20

Passed

Description

87%Weight 40%Scale 1-5

Based on the skill's description, can an agent find and select it at the right time? Clear, specific descriptions lead to better discovery.

A strong, third-person description that answers both what and when with concrete, app-specific triggers and a comprehensive feature enumeration. The only soft spots are slightly generic action verbs and the absence of synonyms/file extensions in the trigger list.

DimensionReasoningScore

Specificity

Names the domain ('WorldMonitor dashboard') and concrete action modes ('existing browser tests', 'scoped manual drive') plus a comprehensive list of feature areas ('panels, map layers, settings, search, country briefs, boot, or dashboard screenshots'), matching the 'several specific actions' anchor with only minor generic-ness in 'verify behavior'.

4 / 5

Completeness

It states the 'what' explicitly ('Verify WorldMonitor dashboard behavior with existing browser tests or a scoped manual drive') and the 'when' with a concrete 'Use for ...' trigger phrase, matching the anchor that requires both answered clearly with concrete triggers.

5 / 5

Trigger Term Quality

The 'Use for' clause enumerates natural feature-name keywords a WorldMonitor user would say (panels, map layers, settings, search, country briefs, boot, screenshots), giving good coverage, but it lacks synonyms and file extensions so it stops short of comprehensive.

4 / 5

Distinctiveness Conflict Risk

The named app 'WorldMonitor' and its app-specific feature triggers give it a clear niche with minimal overlap risk against any other skill.

5 / 5

Total

18

/

20

Passed

Validation

93%

Checks the skill against the spec for correct structure and formatting. All validation checks must pass before discovery and implementation can be scored.

Validation — 15 / 16 Passed

Validation for skill structure

CriteriaDescriptionResult

relative_links

Relative link issues: 7 missing, 4 suspicious

Warning

Total

15

/

16

Passed

Repository
koala73/worldmonitor
Reviewed

Table of Contents

Is this your skill?

If you maintain this skill, you can claim it as your own. Once claimed, you can manage eval scenarios, bundle related skills, attach documentation or rules, and ensure cross-agent compatibility.