CtrlK
BlogDocsLog inGet started
Tessl Logo

debugging-executions

Debug failed or wrong-output workflow executions using executions tools. Load when the user reports execution failures, unexpected node output, empty parameter values after a successful run, or a node showing a red or failed expression error.

68

Quality

86%

Does it follow best practices?

Run evals on this skill

Adds up to 20 points to the overall score

View guide
SecuritybySnyk

Passed

No findings from the security scan

SKILL.md
Quality
Evals
Security

Quality

Content

90%Weight 40%Scale 1-5

Reviews the quality of instructions and guidance provided to agents. Good implementation is clear, handles edge cases, and produces reliable results.

A strong, information-dense body: every section teaches non-obvious n8n execution-tool behavior with exact commands, argument shapes, and safety gates, and it consistently enforces verification before success claims. The main improvements are structural — surfacing an explicit entry-point decision flow and moving the long tool/sub-node edge-case rules into a reference file.

DimensionReasoningScore

Conciseness

The body is dense with non-obvious, product-specific behavior (draft vs. active versionIds, 'unreconstructable-context' tags, run-step refusal rules, mocked-input semantics) and contains essentially no explanation of concepts Claude would already know. Not 4: there is no padded or over-explained passage to trim — each section adds information the model cannot infer; the rhetorical restatements ("sends the message, charges the card, or deletes the row") are doing safety work, not filling space.

5 / 5

Actionability

Concrete, executable tool invocations with full argument shapes throughout: `executions(action="run-step", workflowId, nodeName, reuseExecutionId=...)`, a complete `toolArguments={"query": "login fails", "status": "open"}` example, named actions (`debug`, `get-resolved-node-parameters`, `get-node-output`, `get-as-code`), and specific field names to check (`emptyResolutions`, `failedExpressions`, `workflowVersionId`, `ranThroughNodeNames`). Not 4: the specific examples cover the common cases (failed node, fix confirmation, wrong output on a successful run) with exact commands and semantics, not high-level hints.

5 / 5

Workflow Clarity

Each scenario is sequenced with explicit decision points and verification gates: re-run the failing path before claiming a fix, classify the node (read/transform = safe, write = do not run, unsure = treat as write) before run-step, and confirm a fix only via a new live run after publish — feedback loops are present for the risky operations. Not 5: the overall entry-point flow is implicit — the reader must synthesize the order across sections (when to use `debug` vs `run-step` vs `get-resolved-node-parameters`) rather than following one explicit ordered procedure; not 3: validation and safety checkpoints are explicit, not missing.

4 / 5

Progressive Disclosure

The single file is cleanly sectioned by scenario and the one detail-heavy lookup (trigger inputData shapes) is pushed to a clearly-signaled one-level-deep path: "read ${N8N_WORKSPACE_DIR}/knowledge-base/reference/trigger-input-data-shapes.md". Not 5: no bundle structure exists, and at ~195 lines some specialized material (tool/sub-node/MCP-registry run-step rules) would fit better in a reference file so SKILL.md stays a lean overview; not 3: what is inline is scenario-critical and navigation is easy.

4 / 5

Total

18

/

20

Passed

Description

78%Weight 40%Scale 1-5

Based on the skill's description, can an agent find and select it at the right time? Clear, specific descriptions lead to better discovery.

A well-formed description with an explicit and well-phrased trigger clause; the main weakness is that the 'what' stays at the generic 'debug' level without naming the concrete capabilities the skill actually provides (resolved-parameter inspection, single-node replay, silent-null detection). Distinctiveness and trigger phrasing are strong.

Suggestions

List 2-3 concrete actions in the what-clause, e.g. 'Inspect resolved node parameters, replay a single node against a failed execution's data, and detect expressions that silently resolve to null.'

Add one or two common user phrasings as triggers, such as 'workflow run failed' or 'node output is empty/undefined', to broaden natural-term coverage.

Mention the successful-run-with-wrong-value case in the what-clause, not only the when-clause, so the skill's scope is clear without reading the triggers.

DimensionReasoningScore

Specificity

The description names the domain and one concrete action — "Debug failed or wrong-output workflow executions using executions tools" — but the action is generic ("debug") and the tool-specific capabilities (inspect resolved parameters, replay a single node, detect silent null resolutions) are not listed. Not 4: it does not enumerate several specific actions like the "Extracts text, fills forms, converts pages" anchor; not 2: it is more concrete than a bare domain mention.

3 / 5

Completeness

Explicitly answers both: what — "Debug failed or wrong-output workflow executions using executions tools"; when — "Load when the user reports execution failures, unexpected node output, empty parameter values after a successful run, or a node showing a red or failed expression error", with concrete trigger phrases. This matches the anchor-5 exemplar structure; a 'Use/Load when...' clause is present and specific, so no cap applies.

5 / 5

Trigger Term Quality

Triggers are natural and varied: "execution failures", "unexpected node output", "empty parameter values after a successful run", "a node showing a red or failed expression error" — phrasings a user would plausibly say verbatim. Not 5: common synonyms like "workflow run failed", "node error", or "expression resolved to empty" variations are absent; not 3: coverage goes well beyond a couple of generic keywords.

4 / 5

Distinctiveness Conflict Risk

The triggers are niche and distinctive ("red or failed expression error", "empty parameter values after a successful run") and unlikely to fire for unrelated skills. Not 5: "Debug failed... workflow executions" could mildly overlap with a general workflow-building or verification skill; not 3: the trigger phrases are far more specific than the 'Works with document files' anchor.

4 / 5

Total

16

/

20

Passed

Validation

93%

Checks the skill against the spec for correct structure and formatting. All validation checks must pass before discovery and implementation can be scored.

Validation — 15 / 16 Passed

Validation for skill structure

CriteriaDescriptionResult

frontmatter_unknown_keys

Unknown frontmatter key(s) found; consider removing or moving to metadata

Warning

Total

15

/

16

Passed

Repository
n8n-io/n8n
Reviewed

Table of Contents

Is this your skill?

If you maintain this skill, you can claim it as your own. Once claimed, you can manage eval scenarios, bundle related skills, attach documentation or rules, and ensure cross-agent compatibility.