CtrlK
BlogDocsLog inGet started
Tessl Logo

integration-test

Run the signalbox end-to-end integration evidence run: build the CLI, drive the hub / LAN mode / remote mode / forwarder / hooks / app through real commands with shellwright, screenshot every step, and assemble a self-contained HTML report. Use when the user says "run the integration test", "/integration-test", or wants an evidence run with a report.

72

Quality

88%

Does it follow best practices?

Run evals on this skill

Adds up to 20 points to the overall score

View guide

SecuritybySnyk

Passed

No findings from the security scan

SKILL.md
Quality
Evals
Security

Quality

Content

85%Weight 40%Scale 1-5

Reviews the quality of instructions and guidance provided to agents. Good implementation is clear, handles edge cases, and produces reliable results.

A highly actionable, well-sequenced integration runbook with strong validation and error-recovery guidance. Its main weakness is progressive disclosure: a large monolithic body with no split into reference files.

Suggestions

Move the long UI-scripting steps (10c dropdown overflow, 10d logs pane, 10e shortcut recorder/migration) into reference files under references/ and link to them from the main body to reduce the inline token load.

Extract the per-step evidence-format spec (meta.json/output.txt/*.png) into a short references/evidence-format.md so the Conventions section stays a pointer.

Tighten a few prose passages that restate the verdict policy per step; a single upfront statement plus per-step exceptions would be leaner.

DimensionReasoningScore

Conciseness

Mostly lean operational prose with non-obvious gotchas (port collisions, stale CLI embed, defaults round-trip) that earn their place, though a few explanatory passages could be tightened, keeping it just below the top anchor.

4 / 5

Actionability

Copy-paste-ready bash blocks with real flags, ports (8399/8410/8420), tokens, and osascript calls cover the common scenarios end to end, matching the fully-executable anchor.

5 / 5

Workflow Clarity

Steps 01–12 are explicitly sequenced with per-step pass/warn/fail verdicts, explicit validation checks ('Check for BUILD SUCCEEDED explicitly', md5 match, embedded-version check) and a teardown/restore step providing clear feedback loops.

5 / 5

Progressive Disclosure

Well-organized with clear section headers, but the ~470-line runbook is monolithic with no bundle reference files; content such as the lengthy app/osascript steps (10c–10e) that could live in separate reference files is all inline.

3 / 5

Total

17

/

20

Passed

Description

92%Weight 40%Scale 1-5

Based on the skill's description, can an agent find and select it at the right time? Clear, specific descriptions lead to better discovery.

A strong, third-person description that names concrete actions and gives explicit trigger phrases for both what and when. Slight room to broaden synonym coverage in the trigger clause.

DimensionReasoningScore

Specificity

Lists multiple concrete actions — 'build the CLI', 'drive the hub / LAN mode / remote mode / forwarder / hooks / app through real commands with shellwright', 'screenshot every step', 'assemble a self-contained HTML report' — giving comprehensive coverage of what the run does.

5 / 5

Completeness

Explicitly answers both what (the build/drive/screenshot/report actions) and when ('Use when the user says ...'), with concrete trigger phrases matching the top anchor.

5 / 5

Trigger Term Quality

Provides concrete natural triggers users would say — 'run the integration test', '/integration-test', 'evidence run with a report' — but synonym coverage is narrow and lacks common variations, so it falls just below the comprehensive anchor.

4 / 5

Distinctiveness Conflict Risk

Occupies a clear niche — the signalbox end-to-end integration evidence run — with distinct triggers and minimal overlap risk with other skills.

5 / 5

Total

19

/

20

Passed

Validation

100%

Checks the skill against the spec for correct structure and formatting. All validation checks must pass before discovery and implementation can be scored.

Validation16 / 16 Passed

Validation for skill structure

No warnings or errors.

Repository
dwmkerr/signalbox
Reviewed

Table of Contents

Is this your skill?

If you maintain this skill, you can claim it as your own. Once claimed, you can manage eval scenarios, bundle related skills, attach documentation or rules, and ensure cross-agent compatibility.