CtrlK
BlogDocsLog inGet started
Tessl Logo

resilient-spreadsheet-workflow

Resilient multi-step workflow for spreadsheet processing when execute_code_sandbox fails, using shell_agent exploration, file-based scripts, and verification steps

62

Quality

78%

Does it follow best practices?

Run evals on this skill

Adds up to 20 points to the overall score

View guide
SecuritybySnyk

Passed

No findings from the security scan

Fix and improve this skill with Tessl

tessl review fix ./benchmarks/gdpval/skills/resilient-spreadsheet-workflow/SKILL.md
SKILL.md
Quality
Evals
Security

Quality

Content

73%Weight 40%Scale 1-5

Reviews the quality of instructions and guidance provided to agents. Good implementation is clear, handles edge cases, and produces reliable results.

The body delivers a clear, validated, executable workflow with strong diagnostic guidance, but carries redundant recap content and minor over-explanation that could be trimmed for token efficiency.

Suggestions

Remove or condense the 'Complete Example' section since it restates the step-by-step workflow verbatim.

Drop explanations of basics Claude already knows, such as what '2>&1' does, keeping only the actionable command.

Move the diagnostic error-patterns table or per-step rationale into a references file to tighten SKILL.md toward an overview.

DimensionReasoningScore

Conciseness

Mostly efficient, but includes unnecessary explanation Claude already knows (e.g. 'The 2>&1 redirection ensures both stdout and stderr are captured together') and a 'Complete Example' section that largely restates the prior steps.

3 / 5

Actionability

Provides executable Python and shell commands (read_csv, run_shell with 2>&1, which python, file, head) with only minor gaps such as placeholder inputs and an awkward Python-wrapping of the write_file tool call.

4 / 5

Workflow Clarity

A clearly sequenced five-step workflow with explicit validation checkpoints (Step 4 isolate errors on failure, Step 5 verify output before reading) and a feedback loop, plus a diagnostic error-patterns table acting as a checklist.

5 / 5

Progressive Disclosure

Well-organized into clear sections with no nested references, but at ~125 lines with no bundle files and a redundant 'Complete Example' section it is not a clean single-overview-plus-references structure.

4 / 5

Total

16

/

20

Passed

Description

71%Weight 40%Scale 1-5

Based on the skill's description, can an agent find and select it at the right time? Clear, specific descriptions lead to better discovery.

The description clearly defines a distinct niche with an explicit trigger and lists concrete approaches, but leans on internal tool jargon and omits common user synonyms like CSV/XLSX.

Suggestions

Add natural trigger terms users would actually say, e.g. 'Use when processing CSV or XLSX files that fail to run in the sandbox'.

Replace internal jargon ('execute_code_sandbox', 'shell_agent') with user-facing language or pair it with plain synonyms.

Name a couple of concrete spreadsheet operations (read, transform, export) to lift specificity toward comprehensive coverage.

DimensionReasoningScore

Specificity

Names the spreadsheet domain plus several concrete approaches ('shell_agent exploration, file-based scripts, and verification steps'), but does not enumerate the actual processing operations (filter, transform, aggregate), leaving minor coverage gaps.

4 / 5

Completeness

Both 'what' (multi-step workflow using exploration, file-based scripts, verification) and 'when' ('when execute_code_sandbox fails') are present and explicit, though the trigger is narrow and technical rather than a natural user phrase.

4 / 5

Trigger Term Quality

'spreadsheet processing' is a relevant natural keyword, but the rest is internal-tool jargon ('execute_code_sandbox', 'shell_agent') that users would not say, and synonyms like CSV, XLSX, or Excel are absent.

3 / 5

Distinctiveness Conflict Risk

A clear niche (resilient fallback for spreadsheet processing) gated by a distinct failure-condition trigger ('when execute_code_sandbox fails') minimizes overlap with general spreadsheet skills.

5 / 5

Total

16

/

20

Passed

Validation

100%

Checks the skill against the spec for correct structure and formatting. All validation checks must pass before discovery and implementation can be scored.

Validation — 16 / 16 Passed

Validation for skill structure

No warnings or errors.

Repository
HKUDS/OpenSpace
Reviewed

Table of Contents

Is this your skill?

If you maintain this skill, you can claim it as your own. Once claimed, you can manage eval scenarios, bundle related skills, attach documentation or rules, and ensure cross-agent compatibility.