CtrlK
BlogDocsLog inGet started
Tessl Logo

agent-v3-performance-engineer

Agent skill for v3-performance-engineer - invoke with $agent-v3-performance-engineer

51

3.20x
Quality

25%

Does it follow best practices?

Impact

96%

3.20x

Average score across 3 eval scenarios

SecuritybySnyk

Passed

No findings from the security scan

Fix and improve this skill with Tessl

tessl review fix ./.agents/skills/agent-v3-performance-engineer/SKILL.md
SKILL.md
Quality
Evals
Security

Quality

Content

32%Weight 40%Scale 1-5

Reviews the quality of instructions and guidance provided to agents. Good implementation is clear, handles edge cases, and produces reliable results.

The body is a sprawling, aspirational spec: it documents aggressive numeric targets and benchmark scaffolding but never gives an executable command or a step-by-step procedure for achieving or validating them. Much of it is redundant decoration (ASCII boxes, emoji hooks, triple-stated targets), and the stray second frontmatter block plus the inlined pseudo-code would be better split into real benchmark scripts. Claude reading this learns the goals but not how to act on them.

Suggestions

Move the benchmark classes into executable scripts/ files (e.g. scripts/benchmarks/*.ts with real setup) and keep SKILL.md as a concise overview linking to them.

Replace the pseudocode with runnable commands (e.g. 'npx agentic-flow@alpha bench --suite startup') including defined dependencies, so the guidance is copy-paste executable.

Collapse the triple-stated target lists (hooks echo, ASCII matrix, checklist) into one table and delete the motivational filler; also remove the duplicate inner frontmatter block, which is dead weight in the body.

DimensionReasoningScore

Conciseness

The ~355-line body restates the same target numbers three times (pre-execution hook echoes, the three ASCII 'Performance Target Matrix' boxes, and the 'Target Achievement Checklist') and pads with motivational filler like 'achieve industry-leading performance improvements' and 'the fastest and most efficient agent orchestration platform'. This is noticeably verbose with several padded sections (anchor 2); the benchmark code gives it some substance, so it is not a 1.

2 / 5

Actionability

There is concrete guidance (performance.now() timing loops, improvement ratios, a 5% regression threshold, target ranges) but the TypeScript benchmark classes are pseudocode, not executable: methods like this.spawn15Agents(), this.sona.adapt(scenario), this.initializeCLI(), and the MetricCollector type are never defined. This matches anchor 3 (some concrete guidance but incomplete; pseudocode instead of executable code).

3 / 5

Workflow Clarity

No ordered procedure exists anywhere in the body — it is a target matrix, code stubs, checklists, and coordination notes, not steps. The 'Success Validation Framework' checklist and regression-detection snippet gesture at validation but are not sequenced into a workflow, matching anchor 2 (rough sequence implied at best, steps poorly defined, validation not operational).

2 / 5

Progressive Disclosure

No bundle files exist (no references/, scripts/, or assets/ directories) and 250+ lines of benchmark code are inlined in SKILL.md — content that clearly belongs in separate script files. Section headers exist, but the structure is a monolithic dump matching anchor 2 (minimal structure; content that belongs in separate files is inlined); it does not reach anchor 3 because there are no references at all to signal.

2 / 5

Total

9

/

20

Passed

Description

17%Weight 40%Scale 1-5

Based on the skill's description, can an agent find and select it at the right time? Clear, specific descriptions lead to better discovery.

The description is a self-referential invocation hint rather than a capability description: it says what to invoke, not what the skill does or when to use it. No concrete actions, no natural trigger terms, and no 'Use when...' guidance. Note also the file is malformed — it contains a second frontmatter block with a much richer description ('V3 Performance Engineer for achieving aggressive performance targets...'), but a skill loader parses the outer, near-empty one.

Suggestions

Replace the invocation-only description with third-person concrete actions, e.g. 'Benchmarks and optimizes claude-flow v3 performance: Flash Attention speedup, AgentDB HNSW search latency, memory reduction, and cold-start time.'

Add an explicit trigger clause such as 'Use when the user asks to speed up, benchmark, or profile claude-flow v3, or mentions performance regressions, Flash Attention, or HNSW search.'

Remove the duplicated inner frontmatter block so the richer description becomes the single, parsed `description` field.

DimensionReasoningScore

Specificity

The description ('Agent skill for v3-performance-engineer - invoke with $agent-v3-performance-engineer') names a domain only through the title 'performance-engineer' and states zero concrete actions. It sits between anchor 1 (no concrete actions, entirely vague) and anchor 2 (names the domain but minimal/generic actions) — the domain is named, so 2 rather than 1, but no action verbs appear at all so it cannot reach 3.

2 / 5

Completeness

It offers only a vague 'what' ('Agent skill for v3-performance-engineer') and no 'when' guidance whatsoever; there is no 'Use when...' clause or equivalent. This matches anchor 2 (vague what, no when) and is capped below 3 per the guideline that a missing explicit trigger clause caps completeness at 3 — it is well below that bar.

2 / 5

Trigger Term Quality

The only 'keyword' is the mechanical invocation syntax '$agent-v3-performance-engineer', which no user would naturally say. There are no natural trigger terms such as 'benchmark', 'speedup', 'optimize performance', or 'profiling', matching anchor 1 (no natural keywords; only technical jargon or generic language).

1 / 5

Distinctiveness Conflict Risk

The description is entirely self-referential and provides no capability information, so a model could not distinguish it from other agent/v3 skills based on task content. It is not 'entirely generic' (the agent id is unique), so 2 rather than 1, but overlap risk with any performance-related request is high.

2 / 5

Total

7

/

20

Passed

Validation

100%

Checks the skill against the spec for correct structure and formatting. All validation checks must pass before discovery and implementation can be scored.

Validation — 16 / 16 Passed

Validation for skill structure

No warnings or errors.

Repository
ruvnet/ruflo
Reviewed

Table of Contents

Is this your skill?

If you maintain this skill, you can claim it as your own. Once claimed, you can manage eval scenarios, bundle related skills, attach documentation or rules, and ensure cross-agent compatibility.