CtrlK
BlogDocsLog inGet started
Tessl Logo

testing

Skill: testing

57

1.53x
Quality

39%

Does it follow best practices?

Impact

92%

1.53x

Average score across 3 eval scenarios

SecuritybySnyk

Passed

No findings from the security scan

Fix and improve this skill with Tessl

tessl review fix ./.agents/skills/testing/SKILL.md
SKILL.md
Quality
Evals
Security

Quality

Content

71%Weight 40%Scale 1-5

Reviews the quality of instructions and guidance provided to agents. Good implementation is clear, handles edge cases, and produces reliable results.

The body is a strong, terse, opinionated ruleset: nearly every line is an executable instruction with concrete commands, paths, and thresholds, and the coverage workflow is well sequenced with validation gates. Its main weaknesses are content that belongs in reference files (package rules, allowlists) being inlined in SKILL.md, and verbatim duplication between the seam-selection, core-package, and quick-reference sections.

Suggestions

Move the package-specific rules (`autoformat`, `markdown`, `ai/streaming`, `slate`, `docx`) and the `createPlateEditor`/`__tests__` allowlists into a one-level-deep reference file (e.g. references/package-rules.md) and keep SKILL.md as the overview plus core rules.

De-duplicate the `createSlateEditor` bullet list (Seam Selection vs the `core` section) and the two-lanes/lcov/package-locality rules restated in 'Quick Reference' so each rule is stated once.

DimensionReasoningScore

Conciseness

The body is dense, imperative, and assumes competence — no basic-concept explanations — with concrete thresholds like '60ms/test or 120ms/file' and exact paths ('tooling/config/test-suites.mjs'). It falls short of a 5 because of genuine duplication: the `createSlateEditor` bullet list appears verbatim in both 'Seam Selection' and the `core` package section, and the 'Quick Reference' section restates rules already stated above (two lanes, lcov over Bun summary, package-locality).

4 / 5

Actionability

Guidance is concrete and executable: exact commands ('pnpm test:slowest -- --top 25 --rerun-each 3', 'bun run test:profile'), exact file-naming conventions ('*.spec.ts[x]' vs '*.slow.ts[x]'), exact paths, and numeric coverage-pass thresholds ('>= 6', then '>= 5'). It is not a 5 because some instructions are named but never shown — e.g. 'table-drive repeated node-type cases', 'extract one editor helper', and the hyperscript fixture style have no example snippet, leaving the reader to invent the pattern.

4 / 5

Workflow Clarity

The coverage workflow is explicitly sequenced with passes (>= 6 pass, rerun fresh lcov, >= 5 pass, architecture-safety pass, then explicit stop criteria for wrappers/crumbs/sludge) and includes validation checkpoints and feedback loops (profile -> rename to *.slow.ts -> pnpm test:slowest hard gate). It falls short of a 5 because the workflow is distributed across four sections ('Testing Goal', 'Coverage Strategy', 'Cleanup Heuristics', 'Quick Reference') rather than presented as one coherent ordered procedure.

4 / 5

Progressive Disclosure

Sectioning is clean and there are no dangling or nested references, but the entire rulebook — including package-specific rules (`autoformat`, `markdown`, `slate`, `docx`), allowlists, and the upstream slate-react mining list — lives inline in one 250-line SKILL.md with no reference files at all. That matches 'content that should be separate is inline'; it is above a 2 only because the sections are well-organized and the in-body file paths (e.g. 'tooling/config/bunTestSetup.ts', 'apps/www/src/__tests__/package-integration') are real and clearly signaled.

3 / 5

Total

15

/

20

Passed

Description

7%Weight 40%Scale 1-5

Based on the skill's description, can an agent find and select it at the right time? Clear, specific descriptions lead to better discovery.

The frontmatter description is a placeholder-grade failure. It communicates nothing about the skill's actual scope (a Bun/Plate/Slate monorepo testing strategy with coverage passes and seam selection), so it would never trigger reliably and would collide with any other testing skill.

Suggestions

Replace 'Skill: testing' with a concrete capability summary, e.g. 'Layers a Bun-based test suite into unit, plugin-contract, and golden serializer tests; runs coverage passes ranked by lcov file score and picks the smallest editor seam (createEditor -> createSlateEditor -> createPlateEditor)'.

Add an explicit trigger clause: 'Use when writing or refactoring tests, choosing test seams for Plate/Slate packages, deciding fast-lane vs slow-lane placement, or planning coverage work in this monorepo.'

Include natural trigger keywords such as 'spec', 'test coverage', 'hotspot', 'slow lane', and 'bun test' so the skill surfaces for real user phrasings.

DimensionReasoningScore

Specificity

The description is only 'Skill: testing' — it names no concrete action or capability whatsoever (no layering strategy, no coverage workflow, no commands), which matches the 'entirely vague; no concrete actions' anchor. It cannot score 2 because even the domain-naming-with-minimal-actions example ('Processes PDF files') describes an action verb plus target.

1 / 5

Completeness

Neither 'what does this skill do' nor 'when to use it' is answered in any form — 'Skill: testing' is even vaguer than the anchor example 'Helps with documents'. The missing 'Use when...' clause alone would cap this at 3; here nothing is stated at all.

1 / 5

Trigger Term Quality

The single keyword 'testing' is one generic term with no natural variations users would say when they need this skill (e.g. 'write tests', 'test coverage', 'hotspot', 'bun test', 'spec'), matching the 'one or two generic keywords; missing the natural phrases' anchor. It is above a 1 only because 'testing' is a real natural-language term rather than pure jargon.

2 / 5

Distinctiveness Conflict Risk

The description is entirely generic — 'testing' would match virtually any testing-related request, giving maximal conflict risk with any other test-related skill. It cannot score 2 because even the 'very broad' anchor example at least names a file category.

1 / 5

Total

5

/

20

Passed

Validation

81%

Checks the skill against the spec for correct structure and formatting. All validation checks must pass before discovery and implementation can be scored.

Validation — 13 / 16 Passed

Validation for skill structure

CriteriaDescriptionResult

description_field

'description' is very short (14 chars), consider making it more detailed

Warning

metadata_version

'metadata.version' is missing

Warning

metadata_field

'metadata' should map string keys to string values

Warning

Total

13

/

16

Passed

Repository
udecode/plate
Reviewed

Table of Contents

Is this your skill?

If you maintain this skill, you can claim it as your own. Once claimed, you can manage eval scenarios, bundle related skills, attach documentation or rules, and ensure cross-agent compatibility.