CtrlK
BlogDocsLog inGet started
Tessl Logo

grove-setup

Set up a local Grove environment for running code example tests. Use when the user asks to "set up Grove", "configure the test suite", "get started with Grove", "set up my environment", "I need to run Grove tests", or is working with Grove for the first time and needs prerequisites configured. Checks for required tools, installs dependencies, guides .env file setup, verifies MongoDB connectivity, and checks sample data availability.

75

Quality

94%

Does it follow best practices?

Run evals on this skill

Adds up to 20 points to the overall score

View guide
SecuritybySnyk

—

The risk profile of this skill

SKILL.md
Quality
Evals
Security

Quality

Content

88%Weight 40%Scale 1-5

Reviews the quality of instructions and guidance provided to agents. Good implementation is clear, handles edge cases, and produces reliable results.

A highly actionable, well-sequenced setup guide: every step carries executable commands, validation checkpoints, and failure-recovery guidance. The only real weaknesses are mild verbosity in the Step 0 handoff section and the absence of any progressive-disclosure split despite the file's length.

Suggestions

Trim Step 0's blockquoted example messages to single-line templates or move the full version/shape-check failure messages into a short reference file, cutting ~30 lines of the skill's least-used path.

Consider moving the per-language install commands and smoke-test tables (Steps 3 and 7) into a references/ file keyed by language, keeping SKILL.md to the decision flow and loading only the target language's details.

Tighten the handoff JSON envelopes by documenting the required `context` fields in a compact table instead of full envelope blocks per trigger.

DimensionReasoningScore

Conciseness

The body is dense and operational — tables of exact commands, per-language install snippets, and edge cases — with essentially no explanation of concepts Claude already knows. It falls short of the score-5 'every token earns its place' bar because Step 0's handoff section is padded with long blockquoted example messages and restates schema details ('Example message: > Found a Grove extension handoff...') that could be trimmed to one line each. It clearly exceeds score 3, which would require noticeably unnecessary explanation, not just trimmable verbosity.

4 / 5

Actionability

Nearly every instruction is copy-paste ready: exact install commands per language ('cd code-example-tests/javascript/driver && npm install'), exact smoke-test commands per suite in a table, executable mongosh eval scripts, and a fallback-version table. This matches the score-5 anchor ('Fully executable; copy-paste ready code or commands; specific examples cover the common cases'); score 4 would imply minor gaps in executability, and none are evident.

5 / 5

Workflow Clarity

Steps 0–8 are explicitly sequenced with validation checkpoints and feedback loops: 'Do not proceed past this step until the user confirms', the Step 5 connectivity check with category-based failure diagnosis, the Step 7 smoke test, and the Edge Cases section mapping failures to recovery actions. This matches the score-5 anchor ('Clear sequence with explicit validation steps; feedback loops for error recovery'); it does not fit score 4, whose 'minor validation gaps' are absent.

5 / 5

Progressive Disclosure

The single file is well-structured with clear step headings, tables, and a coherent overview flow, and the branching design (language tables consulted at decision points) is appropriate for inline delivery. It does not reach score 5 because there are no well-signaled separate reference files at all — ~400 lines including per-language setup variants and the multi-trigger handoff schemas that could plausibly live in one-level-deep references; it does not fall to score 3 because the content that is inline genuinely needs to be, and organization is strong, not merely 'could be better organized'.

4 / 5

Total

18

/

20

Passed

Description

100%Weight 40%Scale 1-5

Based on the skill's description, can an agent find and select it at the right time? Clear, specific descriptions lead to better discovery.

The description is exemplary: third-person voice, concrete enumerated capabilities, and an explicit 'Use when' clause with five natural quoted trigger phrases plus a first-time-user condition. It cleanly satisfies the what/when requirements with no padding.

DimensionReasoningScore

Specificity

The description enumerates five concrete actions — 'Checks for required tools, installs dependencies, guides .env file setup, verifies MongoDB connectivity, and checks sample data availability' — which is comprehensive coverage of the skill's capabilities with no vague filler. This matches the score-5 anchor ('Lists multiple specific concrete actions; comprehensive coverage') and exceeds score 4, which would leave minor coverage gaps.

5 / 5

Completeness

Both questions are explicitly answered: 'what' via the five concrete actions and 'when' via 'Use when the user asks to... or is working with Grove for the first time and needs prerequisites configured.' This mirrors the score-5 exemplar's structure (what + 'Use when' with concrete trigger phrases); score 4 would require the 'when' to be less explicit, which it is not.

5 / 5

Trigger Term Quality

It includes five quoted natural phrases users would actually say: '"set up Grove"', '"configure the test suite"', '"get started with Grove"', '"set up my environment"', '"I need to run Grove tests"', plus the first-time-user condition. This is comprehensive natural-term coverage including synonyms, matching the score-5 anchor and exceeding the 'good coverage, a few missing' bar of score 4.

5 / 5

Distinctiveness Conflict Risk

'Grove' is a distinct project-specific noun appearing in every trigger phrase, giving the skill a clear niche with minimal conflict risk against generic setup skills. It matches the score-5 anchor ('Clear niche with distinct triggers'); it does not fall to score 4 because there is no meaningful overlap with closely related skills implied by the description.

5 / 5

Total

20

/

20

Passed

Validation

93%

Checks the skill against the spec for correct structure and formatting. All validation checks must pass before discovery and implementation can be scored.

Validation — 15 / 16 Passed

Validation for skill structure

CriteriaDescriptionResult

frontmatter_unknown_keys

Unknown frontmatter key(s) found; consider removing or moving to metadata

Warning

Total

15

/

16

Passed

Repository
mongodb/docs
Reviewed

Table of Contents

Is this your skill?

If you maintain this skill, you can claim it as your own. Once claimed, you can manage eval scenarios, bundle related skills, attach documentation or rules, and ensure cross-agent compatibility.