CtrlK
BlogDocsLog inGet started
Tessl Logo

add-product-e2e-eval

Add or extend a Paperclip full-stack runner E2E workflow, fixture, matcher, or report evidence path for local or Daytona execution.

60

Quality

71%

Does it follow best practices?

Run evals on this skill

Adds up to 20 points to the overall score

View guide

SecuritybySnyk

Passed

No findings from the security scan

Fix and improve this skill with Tessl

tessl review fix ./.agents/skills/add-product-e2e-eval/SKILL.md
SKILL.md
Quality
Evals
Security

Quality

Content

75%Weight 40%Scale 1-5

Reviews the quality of instructions and guidance provided to agents. Good implementation is clear, handles edge cases, and produces reliable results.

The body is a lean, actionable instruction set with concrete commands, file paths, and validation checkpoints, well-supported by clearly signaled one-level references to authoritative repo docs; the main improvement would be adding section headers for easier navigation.

DimensionReasoningScore

Conciseness

The body is information-dense with non-obvious operational detail and does not re-explain concepts Claude already knows; a few long packed sentences could be trimmed, matching the 'efficient; minor over-explanation' anchor.

4 / 5

Actionability

Concrete executable commands (pnpm test:e2e:runner --list, typecheck, unit) and specific file paths (catalog.ts, fixture-registry.ts, everyday-cases.ts) are provided, mixed with some policy prose rather than fully copy-paste-ready coverage.

4 / 5

Workflow Clarity

A clear paragraph-ordered sequence runs from locating the repo through registration, calibration, discovery confirmation, credential-free checks, and failure attribution, with validation checkpoints present though not as an explicit numbered validate-fix-retry loop.

4 / 5

Progressive Disclosure

Content is appropriately split with one-level-deep, clearly signaled references to authoritative repo docs (README.md, FIXTURES.md, SECURITY.md, EVERYDAY-WORKFLOWS.md, doc/evals.md); the minor gap is the body's own lack of section headers.

4 / 5

Total

16

/

20

Passed

Description

67%Weight 40%Scale 1-5

Based on the skill's description, can an agent find and select it at the right time? Clear, specific descriptions lead to better discovery.

The description is specific and distinctive with concrete actions and a clear niche, but it lacks an explicit 'Use when...' trigger clause and relies on internal product vocabulary rather than natural user phrases, capping completeness and trigger-term quality.

Suggestions

Add an explicit 'Use when...' clause (e.g., 'Use when adding or extending Product E2E evals, fixtures, matchers, or report evidence for the Paperclip runner') to lift completeness above 3.

Include natural trigger phrases or synonyms a user might actually say (e.g., 'end-to-end test', 'eval fixture', 'Daytona run') alongside the product jargon to improve trigger-term quality.

DimensionReasoningScore

Specificity

Names multiple concrete actions — 'workflow, fixture, matcher, or report evidence path' — plus two execution modes ('local or Daytona execution'), giving comprehensive coverage of the skill's scope.

5 / 5

Completeness

The 'what' is clear, but there is no explicit 'Use when...' trigger clause, so per the judging guideline completeness is capped at 3 even though 'when' is weakly implied.

3 / 5

Trigger Term Quality

Relevant product keywords are present (E2E, workflow, fixture, matcher, Daytona) but they are internal product jargon with no natural synonyms or file extensions a user would vary, matching the 'some relevant keywords but missing common variations' anchor.

3 / 5

Distinctiveness Conflict Risk

The description carves a clear Paperclip-specific niche (full-stack runner E2E, local/Daytona) with distinct triggers and minimal overlap with other skills.

5 / 5

Total

16

/

20

Passed

Validation

93%

Checks the skill against the spec for correct structure and formatting. All validation checks must pass before discovery and implementation can be scored.

Validation — 15 / 16 Passed

Validation for skill structure

CriteriaDescriptionResult

relative_links

Relative link issues: 1 suspicious

Warning

Total

15

/

16

Passed

Repository
paperclipai/paperclip
Reviewed

Table of Contents

Is this your skill?

If you maintain this skill, you can claim it as your own. Once claimed, you can manage eval scenarios, bundle related skills, attach documentation or rules, and ensure cross-agent compatibility.