CtrlK
BlogDocsLog inGet started
Tessl Logo

config-evals

Builds and maintains configuration-based evaluations on a workflow with the eval-config tool. Use when the user asks to set up, add, view, change, or remove an evaluation, score, grade, or judge a workflow's output, or measure answer quality against a test dataset. This is the only eval form Instance AI handles — it does not touch on-canvas evaluation nodes.

79

Quality

100%

Does it follow best practices?

Run evals on this skill

Adds up to 20 points to the overall score

View guide

SecuritybySnyk

Passed

No findings from the security scan

SKILL.md
Quality
Evals
Security

Quality

Content

100%Weight 40%Scale 1-3

Reviews the quality of instructions and guidance provided to agents. Good implementation is clear, handles edge cases, and produces reliable results.

A well-structured instruction-only skill: lean body covering the procedure and sharp edges, with concrete tool/field guidance and a single clearly-signaled reference for worked examples. Workflow is sequenced with explicit guards and a verification step.

DimensionReasoningScore

Conciseness

Dense, operational guidance with no explanations of concepts Claude already knows; every section (node selection, the leading-'=' sharp edge, dataset boundary) earns its place, matching the lean-and-efficient anchor.

3 / 3

Actionability

Names concrete tool actions and every config/metric field, gives copy-paste expressions with correct/wrong examples, and defers full call recipes to the reference; per the scoring notes, absence of code in this instruction-only skill is not penalized because the guidance is actionable.

3 / 3

Workflow Clarity

A numbered six-step Default Procedure with explicit guards (approval-card checkpoint, 'Never invent a dataTableId', the leading-'=' rule) and a closing 'Close with facts' verification step; not 2 because checkpoints are explicit rather than implicit.

3 / 3

Progressive Disclosure

A concise overview body with one well-signaled, one-level-deep reference ('Use references/config-eval-playbook.md for tool-call recipes, worked examples, and output shapes') that is a real file holding appropriately deeper detail.

3 / 3

Total

12

/

12

Passed

Description

100%Weight 40%Scale 1-3

Based on the skill's description, can an agent find and select it at the right time? Clear, specific descriptions lead to better discovery.

A strong description: third-person, specific, with an explicit Use-when trigger clause and a clear boundary that distinguishes it from on-canvas evaluation skills. It covers what, when, and which actions comprehensively without padding.

DimensionReasoningScore

Specificity

Lists many concrete actions — 'set up, add, view, change, or remove… score, grade, or judge… measure answer quality' — matching the multiple-specific-actions anchor; not 2 because coverage is comprehensive rather than partial.

3 / 3

Completeness

Explicitly answers both what ('Builds and maintains configuration-based evaluations… with the eval-config tool') and when (an explicit 'Use when…' clause); the Use-when clause is present so completeness is not capped at 2.

3 / 3

Trigger Term Quality

'Use when the user asks to set up, add, view, change, or remove an evaluation, score, grade, or judge a workflow's output' uses natural verbs a user would actually say; good coverage, matching the anchor-3 example.

3 / 3

Distinctiveness Conflict Risk

Carves a clear niche — 'the only eval form Instance AI handles — it does not touch on-canvas evaluation nodes' — and explicitly excludes the competing form, making conflict unlikely.

3 / 3

Total

12

/

12

Passed

Validation

93%

Checks the skill against the spec for correct structure and formatting. All validation checks must pass before discovery and implementation can be scored.

Validation15 / 16 Passed

Validation for skill structure

CriteriaDescriptionResult

frontmatter_unknown_keys

Unknown frontmatter key(s) found; consider removing or moving to metadata

Warning

Total

15

/

16

Passed

Repository
n8n-io/n8n
Reviewed

Table of Contents

Is this your skill?

If you maintain this skill, you can claim it as your own. Once claimed, you can manage eval scenarios, bundle related skills, attach documentation or rules, and ensure cross-agent compatibility.