CtrlK
BlogDocsLog inGet started
Tessl Logo

auto-perf-optimize

Run agent-driven VS Code performance or memory investigations. Use when asked to launch Code OSS, automate a VS Code scenario, run the Chat memory smoke runner, capture renderer heap snapshots, take workflow screenshots, compare run summaries, or drive a repeatable scenario before heap-snapshot analysis.

71

Quality

88%

Does it follow best practices?

Run evals on this skill

Adds up to 20 points to the overall score

View guide

SecuritybySnyk

Low

Low-risk findings worth noting

The canonical home for this skill is auto-perf-optimize in microsoft/vscode

SKILL.md
Quality
Evals
Security

Quality

Content

88%Weight 40%Scale 1-5

Reviews the quality of instructions and guidance provided to agents. Good implementation is clear, handles edge cases, and produces reliable results.

A highly actionable, well-sequenced operational skill: real commands, explicit validation and feedback loops, strong safety guidance, and a clear handoff boundary to heap-snapshot-analysis. The main improvements are trimming repeated guidance (scratchpad convention, throwaway-workspace warnings) and splitting the diagnosis/fix playbook into a reference file.

DimensionReasoningScore

Conciseness

Nearly every token carries repo-specific knowledge Claude cannot know (runner flags, profile seeding rules, import-path depth, snapshot-label conventions), but there is real repetition — the dated-scratchpad-subfolder convention is explained in two sections and the throwaway-workspace warning appears three times. This fits 'efficient; minor instances of over-explanation that could be trimmed' rather than the mostly-efficient level below, since the padding is minor relative to the substance.

4 / 5

Actionability

Every workflow is backed by copy-paste-ready commands with real flags and paths (e.g., 'node .github/skills/auto-perf-optimize/scripts/chat-memory-smoke.mts --iterations 8 --heap-snapshot-label 03-iteration-01 …') plus a complete, runnable comparison script, matching the 'fully executable, covers common cases' top anchor; not 4 because no pseudocode or missing-key-details gaps exist.

5 / 5

Workflow Clarity

'The Story' gives an explicit 8-step sequence with validation checkpoints and feedback loops: validate with '--no-heap-snapshots' first, a 'Verify the run'/'Verify Before Analyzing' summary.json and screenshot checklist before opening snapshots, and a fix → rerun → compare-like-for-like verification loop. This matches the top anchor including error-recovery guidance (diagnose stuck runs via incremental summary.json and last screenshot).

5 / 5

Progressive Disclosure

All three linked script files exist and links are one level deep with clear signaling, and section headers make the ~280-line body easy to navigate — but no reference files exist, so sections like 'Root-Cause, Don't Treat Symptoms' and the Playwright watch-loop checkpoints are inlined in one file when they could be split out. 'Good structure; most content appropriately placed; minor organization gaps' fits better than the score-5 clear-split anchor, and better than 3 since references that do exist are clear and navigation is easy.

4 / 5

Total

18

/

20

Passed

Description

88%Weight 40%Scale 1-5

Based on the skill's description, can an agent find and select it at the right time? Clear, specific descriptions lead to better discovery.

A strong description: third-person, concrete, and comprehensive, with an explicit 'Use when' trigger clause enumerating specific user requests. The only weaknesses are a few missing natural synonyms (e.g., 'memory leak', 'profile') and minor overlap with the heap-snapshot-analysis sibling skill.

DimensionReasoningScore

Specificity

Names the domain and multiple concrete actions — 'launch Code OSS', 'run the Chat memory smoke runner', 'capture renderer heap snapshots', 'take workflow screenshots', 'compare run summaries' — with comprehensive coverage and no meaningful gaps, matching the top anchor rather than the 'minor gaps' level below.

5 / 5

Completeness

Explicitly answers both: what it does ('Run agent-driven VS Code performance or memory investigations') and when to use it ('Use when asked to launch Code OSS, automate a VS Code scenario, …') with concrete trigger phrases, exactly matching the top anchor; not 4 because the 'when' clause is fully explicit, not improvable.

5 / 5

Trigger Term Quality

Good natural-term coverage ('launch VS Code', 'memory', 'heap snapshots', 'Chat memory smoke runner', 'workflow screenshots') but misses common variations users would say such as 'memory leak', 'profile', 'perf', or the '.heapsnapshot' extension, so it falls between the good and comprehensive anchors.

4 / 5

Distinctiveness Conflict Risk

A clear VS Code/Code OSS memory-investigation niche with distinct triggers, but 'capture renderer heap snapshots' overlaps the sibling heap-snapshot-analysis skill — partially mitigated by 'before heap-snapshot analysis' — so 'mostly distinct; minor overlap risk' fits better than the minimal-conflict top anchor.

4 / 5

Total

18

/

20

Passed

Validation

81%

Checks the skill against the spec for correct structure and formatting. All validation checks must pass before discovery and implementation can be scored.

Validation — 13 / 16 Passed

Validation for skill structure

CriteriaDescriptionResult

metadata_version

'metadata.version' is missing

Warning

relative_links

Relative link issues: 2 missing

Warning

referenced_paths_exist

Referenced path issues: 2 missing

Warning

Total

13

/

16

Passed

Repository
posit-dev/positron
Reviewed

Table of Contents

Is this your skill?

If you maintain this skill, you can claim it as your own. Once claimed, you can manage eval scenarios, bundle related skills, attach documentation or rules, and ensure cross-agent compatibility.