CtrlK
BlogDocsLog inGet started
Tessl Logo

antithesis-triage

Triage Antithesis test reports to understand what happened in a run: look up runs, check status, investigate failed properties (assertions), view metadata, download logs, inspect findings, and examine environmental details. Load after a run completes or when investigating a failure.

75

Quality

92%

Does it follow best practices?

Run evals on this skill

Adds up to 20 points to the overall score

View guide

SecuritybySnyk

Low

Low-risk findings worth noting

SKILL.md
Quality
Evals
Security

Quality

Content

92%Weight 40%Scale 1-5

Reviews the quality of instructions and guidance provided to agents. Good implementation is clear, handles edge cases, and produces reliable results.

The content is highly actionable and well-structured, with clear workflows, validation checkpoints, and excellent progressive disclosure via real reference files. Only minor conciseness trims would bring it to the top anchor across the board.

DimensionReasoningScore

Conciseness

The body is mostly lean and assumes Claude's competence, with concrete commands throughout, but a few spots (e.g. the 'This validates API connectivity...' gloss and the extended follow-up-skills catalog) could be trimmed.

4 / 5

Actionability

Provides copy-paste-ready commands like 'snouty doctor --json', 'snouty runs --json events ${RUN_ID} ${PROPERTY_NAME}', and concrete jq queries, covering the common triage cases.

5 / 5

Workflow Clarity

Workflows are clearly sequenced with explicit validation checkpoints — the preflight gates on top-level 'ok' and api_key status, and the triage/incomplete-run sections give ordered steps with feedback for error cases.

5 / 5

Progressive Disclosure

SKILL.md is a clear overview with five well-signaled one-level-deep references (run-discovery, run-info, properties, logs, instrumentation), each named at the point of need, with explicit guidance not to load them all up front; all referenced files exist.

5 / 5

Total

19

/

20

Passed

Description

92%Weight 40%Scale 1-5

Based on the skill's description, can an agent find and select it at the right time? Clear, specific descriptions lead to better discovery.

The description is specific, complete, and clearly distinct, with explicit 'what' and 'when' guidance. It could add a few more natural synonyms/shorthand terms users say to push trigger coverage to the top anchor.

DimensionReasoningScore

Specificity

Lists multiple concrete actions — 'look up runs, check status, investigate failed properties (assertions), view metadata, download logs, inspect findings, and examine environmental details' — giving comprehensive coverage of the skill's capabilities.

5 / 5

Completeness

Clearly states both what it does (the enumerated triage actions) and when to use it ('Load after a run completes or when investigating a failure'), with an explicit trigger clause.

5 / 5

Trigger Term Quality

Includes natural user-facing terms like 'test reports', 'failed properties', 'investigating a failure', and 'run', but is missing common synonyms and shorthand variants users might say.

4 / 5

Distinctiveness Conflict Risk

Occupies a clear Antithesis-specific niche with distinct triggers (test reports, runs, properties), so overlap with unrelated skills is minimal.

5 / 5

Total

19

/

20

Passed

Validation

100%

Checks the skill against the spec for correct structure and formatting. All validation checks must pass before discovery and implementation can be scored.

Validation16 / 16 Passed

Validation for skill structure

No warnings or errors.

Repository
antithesishq/antithesis-skills
Reviewed

Table of Contents

Is this your skill?

If you maintain this skill, you can claim it as your own. Once claimed, you can manage eval scenarios, bundle related skills, attach documentation or rules, and ensure cross-agent compatibility.