CtrlK
BlogDocsLog inGet started
Tessl Logo

cursor-agent-supervisor

Offloading tasks with a well-defined scope to sub-agents, for instance to use a sub-agent to implement a set of specs. Use this skill whenever a task should not need a broad knowledge of the whole project

79

1.59x
Quality

71%

Does it follow best practices?

Impact

91%

1.59x

Average score across 3 eval scenarios

SecuritybySnyk

Passed

No findings from the security scan

Fix and improve this skill with Tessl

tessl review fix ./cursor-agent-supervisor/SKILL.md
SKILL.md
Quality
Evals
Security

Quality

Content

81%Weight 40%Scale 1-5

Reviews the quality of instructions and guidance provided to agents. Good implementation is clear, handles edge cases, and produces reliable results.

The content is a strong, lifecycle-complete guide with executable commands, explicit verification steps, and error-recovery loops — workflow clarity is excellent. Weaknesses are minor: some repetitive phrasing, the absence of a filled-in example invocation, and a template/failure playbook inlined in SKILL.md that could be split into reference files.

Suggestions

Trim repetition and rhetorical emphasis — e.g., collapse 'hack on code, run build/tests, hack on code, etc.' and the 'YOU are in charge of bookkeeping...' lines into one directive sentence.

Add one filled-in example of a real sub-agent invocation (model, a sample task description, expected reply) so the commands and template can be used verbatim.

Move the task-description template and the 'When Things Go Wrong' playbook into a reference file (e.g., references/task-template.md) and keep SKILL.md as a lean overview with one-level-deep pointers.

DimensionReasoningScore

Conciseness

The body is mostly lean — concrete commands, a compact Do/Don't list, and a template — but has minor trims available: 'hack on code, run build/tests, hack on code, etc.' repeats 'hack on code', and lines like 'YOU are in charge of bookkeeping, not them. YOU have the big picture, they don't.' add emphasis over information. This fits 'efficient; minor instances of over-explanation that could be trimmed' rather than the every-token-earns-its-place 5 anchor.

4 / 5

Actionability

It gives executable commands ('cursor-agent --print --model <model-name> create-chat', '--resume <conversation-uuid>'), a concrete default model ('sonnet-4.5'), a model-discovery trick, an explicit timeout setting (600000), and a structured task template. The gap versus the 5 anchor is that no filled-in worked example of a real invocation/task is shown, and the commands still require the user to substitute placeholders without an instance of doing so.

4 / 5

Workflow Clarity

The sections follow the full lifecycle — create chat, model selection, invocation with timeout, setup, giving instructions, post-completion verification, and failure recovery — with explicit validation ('Always verify the sub-agent's work: check the diff, review code, run tests/lint') and feedback loops for error recovery ('Read the output, fix any blocking issues, retry with more context', 'Continue the work yourself or spawn another sub-agent with clarified instructions'). This matches the anchor with explicit validation steps and retry loops, not merely 'most checkpoints present'.

5 / 5

Progressive Disclosure

The body (~95 lines) has no bundle files and is well-sectioned with clear headers covering each lifecycle phase. However, it exceeds the under-50-lines simple-skill threshold, and content such as the full task-description template and the failure-mode playbook could plausibly live in a reference file, leaving SKILL.md a leaner overview. This fits 'good structure; most content appropriately placed; minor organization gaps' rather than the ideally-split 5 anchor.

4 / 5

Total

17

/

20

Passed

Description

62%Weight 40%Scale 1-5

Based on the skill's description, can an agent find and select it at the right time? Clear, specific descriptions lead to better discovery.

The description answers both what and when with an explicit 'Use this skill whenever...' clause and names a concrete example use case, but its trigger vocabulary is thin and its capability list is minimal. Adding natural synonyms (delegate, fan out, parallelize) and enumerating a couple more concrete capabilities would lift the two 3-scores.

Suggestions

Add natural trigger synonyms users would actually say, such as 'delegate', 'fan out work', or 'run tasks in parallel with sub-agents', to broaden trigger-term coverage.

Enumerate 1-2 more concrete capabilities (e.g., 'spawn sub-agents, give them scoped tasks, and verify their results') to strengthen the specificity of the 'what' clause.

Rephrase the 'when' clause positively with concrete triggers, e.g., 'Use when a task has a well-defined scope, such as implementing a spec or fixing an isolated bug, and does not require whole-project context.'

DimensionReasoningScore

Specificity

The description names the domain ('Offloading tasks with a well-defined scope to sub-agents') and one concrete action ('use a sub-agent to implement a set of specs'), matching the anchor for naming the domain plus 1-2 concrete actions without comprehensive coverage. It does not list several specific actions (spawn/monitor/verify sub-agents), so it is not a 4.

3 / 5

Completeness

It explicitly answers both: what ('Offloading tasks with a well-defined scope to sub-agents... implement a set of specs') and when ('Use this skill whenever a task should not need a broad knowledge of the whole project'). The 'when' clause is present but phrased negatively and lacks concrete trigger phrases, matching the anchor where both exist but the 'when' could be more explicit — not the 5 anchor's concrete trigger phrasing.

4 / 5

Trigger Term Quality

It includes relevant keywords like 'offloading tasks', 'sub-agents', 'implement a set of specs', and 'broad knowledge of the whole project', but misses common natural variations users would say such as 'delegate', 'delegation', 'fan out', or 'parallelize work'. This matches the 'some relevant keywords but missing common variations or synonyms' anchor rather than the good-coverage anchor.

3 / 5

Distinctiveness Conflict Risk

Sub-agent supervision with a well-defined-scope, no-broad-context-needed framing carves a fairly distinct niche with low overlap risk against typical file-format or workflow skills. Minor overlap risk remains with general delegation/coordination skills, fitting 'mostly distinct; minor overlap risk' rather than the minimal-conflict 5 anchor.

4 / 5

Total

14

/

20

Passed

Validation

100%

Checks the skill against the spec for correct structure and formatting. All validation checks must pass before discovery and implementation can be scored.

Validation — 16 / 16 Passed

Validation for skill structure

No warnings or errors.

Repository
YPares/agent-skills
Reviewed

Table of Contents

Is this your skill?

If you maintain this skill, you can claim it as your own. Once claimed, you can manage eval scenarios, bundle related skills, attach documentation or rules, and ensure cross-agent compatibility.