CtrlK
BlogDocsLog inGet started
Tessl Logo

agent-test-long-runner

Agent skill for test-long-runner - invoke with $agent-test-long-runner

54

0.98x
Quality

30%

Does it follow best practices?

Impact

96%

0.98x

Average score across 3 eval scenarios

SecuritybySnyk

Passed

No findings from the security scan

Fix and improve this skill with Tessl

tessl review fix ./.agents/skills/agent-test-long-runner/SKILL.md
SKILL.md
Quality
Evals
Security

Quality

Content

32%Weight 40%Scale 1-5

Reviews the quality of instructions and guidance provided to agents. Good implementation is clear, handles edge cases, and produces reliable results.

The body is generic agent-persona boilerplate: capability labels, work-style attitudes, and a vague output format with no concrete guidance, commands, or examples. It tells Claude how to behave rather than what to do, so it would not meaningfully change behavior on any specific task.

Suggestions

Replace the attitude-based 'Instructions' with a concrete task workflow (e.g. scope the task, plan, execute in checkpoints, verify against requirements) with explicit validation steps.

Add executable specifics for at least one or two of the named use cases — a sample analysis structure, a report template, or a verification command — instead of 'Code examples where applicable'.

Delete the duplicate embedded '---' frontmatter block at the top of the body and the filler closing line; fold any real signal (long-running, 30+ minute tasks) into the description instead.

DimensionReasoningScore

Conciseness

The body is short and does not explain concepts Claude already knows, but sections like 'Remember: You have plenty of time to do thorough, high-quality work!' and the attitude list are filler that earns no tokens. This matches anchor 3 (mostly efficient, some unnecessary content) rather than 4, where every section would carry actionable weight.

3 / 5

Actionability

There is no code, command, or concrete instruction anywhere — bullets like 'Complex Analysis: Deep dive into codebases, documentation, and systems' and 'Take Your Time: Don't rush - quality over speed' only describe attitudes. This squarely matches anchor 1 (entirely vague or abstract; only describes rather than instructs).

1 / 5

Workflow Clarity

The 'Instructions' section is numbered but lists dispositions ('Be Thorough', 'Document Everything') rather than task steps, and there are no validation checkpoints or feedback loops. This matches anchor 2 (rough sequence present but steps poorly defined); not 1 because a numbered list and an explicit output-format section do exist.

2 / 5

Progressive Disclosure

The skill is under 50 lines with no external references needed, and section headers (Capabilities, Instructions, Output Format, Example Use Cases) are clearly organized. It falls short of 5 because the body opens with a stray duplicate '---' frontmatter block, a structural defect that confuses navigation.

4 / 5

Total

10

/

20

Passed

Description

28%Weight 40%Scale 1-5

Based on the skill's description, can an agent find and select it at the right time? Clear, specific descriptions lead to better discovery.

The description is boilerplate: a label and an invocation hint with no statement of capabilities, use cases, or natural trigger phrases. A user or Claude scanning it could not tell when to invoke this skill or what it does beyond its name.

Suggestions

State 2-3 concrete capabilities in third person, e.g. 'Performs long-running (30+ minute) analysis, research, and system-design tasks that require deep iteration'.

Add an explicit 'Use when...' clause with natural trigger phrases such as 'long-running task', 'deep codebase analysis', or 'comprehensive research report'.

Remove the meta-invocation note ('invoke with $agent-test-long-runner') — it describes plumbing, not the skill's purpose, and wastes the description's limited budget.

DimensionReasoningScore

Specificity

The description names the domain ('test-long-runner') but describes no concrete actions — 'Agent skill for test-long-runner - invoke with $agent-test-long-runner' is only a label plus an invocation note. It matches anchor 2 (domain named, actions minimal/generic) rather than 3, which requires 1-2 stated capabilities.

2 / 5

Completeness

It offers a vague 'what' ('Agent skill for test-long-runner') and no 'when' guidance at all, matching anchor 2 ('Has a vague what and no when'). It is not 3 because the 'what' never states what the skill actually does, and not 1 because a domain and invocation path are at least identified.

2 / 5

Trigger Term Quality

The only keyword is the technical token '$agent-test-long-runner'; there are no natural phrases a user would say such as 'long-running task' or 'deep analysis'. This sits at anchor 2 (one or two keywords, missing the natural phrases) rather than 3, which expects several relevant keywords.

2 / 5

Distinctiveness Conflict Risk

The named-skill trigger '$agent-test-long-runner' is fairly non-colliding, but the boilerplate framing 'Agent skill for X' would match virtually any agent skill. This fits anchor 3 (somewhat specific, could still overlap); not 4 because no distinct trigger phrases beyond the slash-command token are given.

3 / 5

Total

9

/

20

Passed

Validation

100%

Checks the skill against the spec for correct structure and formatting. All validation checks must pass before discovery and implementation can be scored.

Validation — 16 / 16 Passed

Validation for skill structure

No warnings or errors.

Repository
ruvnet/ruflo
Reviewed

Table of Contents

Is this your skill?

If you maintain this skill, you can claim it as your own. Once claimed, you can manage eval scenarios, bundle related skills, attach documentation or rules, and ensure cross-agent compatibility.