CtrlK
BlogDocsLog inGet started
Tessl Logo

testerarmy-cli

Use TesterArmy CLI to create, organize, and run dashboard-managed QA tests. Prefer saved tests, groups, project context, credentials, and remote runs over one-off local prompts. Trigger when defining regression coverage, adding QA flows, or wiring CI checks.

69

Quality

85%

Does it follow best practices?

Run evals on this skill

Adds up to 20 points to the overall score

View guide

SecuritybySnyk

Passed

No findings from the security scan

SKILL.md
Quality
Evals
Security

Quality

Content

82%Weight 40%Scale 1-5

Reviews the quality of instructions and guidance provided to agents. Good implementation is clear, handles edge cases, and produces reliable results.

A well-structured, highly actionable CLI skill: executable commands with concrete payloads cover every workflow, and validation checkpoints (auth check, inspect-before-change, remote --wait) are present. The main refinements are trimming the restated defaults, adding run-failure handling guidance, and moving the payload spec or mobile details into a reference file.

DimensionReasoningScore

Conciseness

Command-first writing with essentially zero explanation of concepts Claude already knows; every section is commands, payloads, or rules. Minor over-repetition keeps it from anchor 5: the "Defaults" list restates the "Modes" bullets, and the Local Prompt section repeats the intro's warning about one-off runs. Clearly above the midpoint, so 4 rather than 3.

4 / 5

Actionability

Copy-paste-ready executable commands throughout — auth, discovery, project/credential/test/group creation with full JSON payloads, run modes, CI, and mobile upload — plus a concrete payload spec and step-type rules. Matches anchor 5: specific examples cover the common cases.

5 / 5

Workflow Clarity

Clear sequence (check auth → discover scope → create context → tests → groups → run/validate) with real checkpoints: "Check auth", "Inspect before changing" before updates, and remote --wait validation. Not 5 because there is no error-recovery guidance for failed runs or auth problems; not 3 because validation checkpoints are explicit rather than missing.

4 / 5

Progressive Disclosure

Good structure with a clearly signaled one-level-deep reference (References table linking references/reporting-template.md, verified to exist), and the body is a reasonable CLI overview rather than a wall of text. Kept at 4 because bulk material — the full JSON payload spec and Mobile App Coverage details — is inlined where a second reference file could carry it. No buried or nested references, so above 3.

4 / 5

Total

17

/

20

Passed

Description

87%Weight 40%Scale 1-5

Based on the skill's description, can an agent find and select it at the right time? Clear, specific descriptions lead to better discovery.

A strong description with explicit what-and-when structure, third-person-style verb phrasing, and a named tool that makes it highly distinguishable. Trigger terms are natural and relevant, though a few common QA synonyms (smoke, E2E, test suite) would round out coverage.

DimensionReasoningScore

Specificity

Names the domain and three concrete actions ("create, organize, and run dashboard-managed QA tests") plus enumerated working objects (tests, groups, credentials, remote runs), but coverage is not fully comprehensive — reporting, mobile uploads, and CI run details appear only in the body. Falls between the 4 and 5 anchors, closer to 4 for minor gaps.

4 / 5

Completeness

Explicitly answers both parts: what ("create, organize, and run dashboard-managed QA tests") and when ("Trigger when defining regression coverage, adding QA flows, or wiring CI checks") with concrete trigger phrases. Matches anchor 5; the when clause is explicit rather than implied, ruling out 4.

5 / 5

Trigger Term Quality

"regression coverage", "QA flows", and "CI checks" are natural trigger phrases, but common synonyms such as smoke tests, E2E tests, or test suites are missing. Good keyword coverage with a few natural terms absent — matches anchor 4, not 5.

4 / 5

Distinctiveness Conflict Risk

Names the specific tool ("TesterArmy CLI") with a clear dashboard-managed QA niche and distinct triggers, giving minimal conflict risk with other skills. Matches anchor 5; the named-tool scope is more distinct than the 'mostly distinct' anchor 4.

5 / 5

Total

18

/

20

Passed

Validation

93%

Checks the skill against the spec for correct structure and formatting. All validation checks must pass before discovery and implementation can be scored.

Validation — 15 / 16 Passed

Validation for skill structure

CriteriaDescriptionResult

metadata_version

'metadata.version' is missing

Warning

Total

15

/

16

Passed

Repository
novuhq/novu
Reviewed

Table of Contents

Is this your skill?

If you maintain this skill, you can claim it as your own. Once claimed, you can manage eval scenarios, bundle related skills, attach documentation or rules, and ensure cross-agent compatibility.