CtrlK
BlogDocsLog inGet started
Tessl Logo

e2e

Write, run, review, or debug rtp2httpd E2E tests and their harness in e2e/ and scripts/run-e2e.sh.

68

Quality

84%

Does it follow best practices?

Run evals on this skill

Adds up to 20 points to the overall score

View guide

SecuritybySnyk

Passed

No findings from the security scan

SKILL.md
Quality
Evals
Security

Quality

Content

93%Weight 40%Scale 1-5

Reviews the quality of instructions and guidance provided to agents. Good implementation is clear, handles edge cases, and produces reliable results.

The body is lean, highly actionable, and well-structured with appropriately signaled one-level-deep references and embedded validation steps. Workflow clarity could be tightened slightly with an explicit numbered checklist for the completion sequence.

Suggestions

Render the Completion section as a short numbered checklist (run affected → ruff check → ruff format --check → collection check) with an explicit "if a check fails, fix and re-run" feedback loop.

Consider noting when to re-run the full suite vs. affected cases in a single explicit rule to remove the mild ambiguity in the closing paragraph.

DimensionReasoningScore

Conciseness

The ~30-line body assumes Claude knows pytest/uv/xdist and contains no concept padding; every line (wrapper defaults, marker location, binary-vs-collection note) earns its place.

5 / 5

Actionability

Copy-paste-ready commands (./scripts/run-e2e.sh …, uv run --group dev ruff check/format) cover the common single-file, serial, parallel, collection, and full-suite cases.

5 / 5

Workflow Clarity

A clear run → choose-detail → completion sequence with explicit validation (ruff check, ruff format --check, collection check) is present, but the steps are prose rather than a numbered checklist with explicit feedback loops.

4 / 5

Progressive Disclosure

The body is a concise overview that points to real one-level-deep references (references/authoring.md, references/troubleshooting.md), each clearly signaled by the condition that triggers reading it.

5 / 5

Total

19

/

20

Passed

Description

75%Weight 40%Scale 1-5

Based on the skill's description, can an agent find and select it at the right time? Clear, specific descriptions lead to better discovery.

The description is specific, actionable, and highly distinct, but lacks an explicit "Use when…" trigger clause, which caps its completeness. Adding a trigger phrase would lift the completeness dimension.

Suggestions

Append an explicit trigger clause, e.g. "Use when writing, running, reviewing, or debugging rtp2httpd E2E tests, or when the user mentions E2E/end-to-end tests or test hangs."

Add natural synonyms such as "end-to-end tests" or "flaky tests" to broaden trigger-term coverage.

DimensionReasoningScore

Specificity

"Write, run, review, or debug" lists multiple concrete actions over "rtp2httpd E2E tests and their harness" with exact paths (e2e/, scripts/run-e2e.sh), giving comprehensive coverage of the skill's tasks.

5 / 5

Completeness

The "what" is clear and concrete, but there is no "Use when…" clause or equivalent explicit trigger guidance, so completeness is capped at 3 per the rubric guidelines.

3 / 5

Trigger Term Quality

Natural terms like "E2E tests", "run", "review", "debug", and "harness" are present, but common synonyms such as "end-to-end" or "flaky tests" are missing.

4 / 5

Distinctiveness Conflict Risk

The rtp2httpd-specific scope with exact directory and script paths carves a clear niche with minimal overlap risk against other skills.

5 / 5

Total

17

/

20

Passed

Validation

93%

Checks the skill against the spec for correct structure and formatting. All validation checks must pass before discovery and implementation can be scored.

Validation15 / 16 Passed

Validation for skill structure

CriteriaDescriptionResult

relative_links

Relative link issues: 1 suspicious

Warning

Total

15

/

16

Passed

Repository
stackia/rtp2httpd
Reviewed

Table of Contents

Is this your skill?

If you maintain this skill, you can claim it as your own. Once claimed, you can manage eval scenarios, bundle related skills, attach documentation or rules, and ensure cross-agent compatibility.