CtrlK
BlogDocsLog inGet started
Tessl Logo

tidb-realtikv-runner

Use when running tests under tests/realtikvtest that require a local TiUP playground lifecycle with strict startup, readiness checks, and cleanup.

87

1.09x
Quality

84%

Does it follow best practices?

Impact

100%

1.09x

Average score across 3 eval scenarios

SecuritybySnyk

Passed

No findings from the security scan

SKILL.md
Quality
Evals
Security

Quality

Content

82%Weight 40%Scale 1-5

Reviews the quality of instructions and guidance provided to agents. Good implementation is clear, handles edge cases, and produces reliable results.

A disciplined, token-efficient checklist body with concrete parameter defaults, test-scoping guidance, and explicit validation checkpoints for readiness and teardown. Its weaknesses are the absence of inline executable commands (deferred to an external doc) and any error-recovery guidance when readiness or teardown fails.

Suggestions

Inline the core commands (playground start invocation and the readiness check) or vendor them into a references/ file inside the bundle, since docs/agents/testing-flow.md is outside the skill and cannot be guaranteed present.

Add a short failure-recovery step, e.g. what to do if the readiness check does not pass within a timeout (port-conflict fallback is mentioned, but no retry/abort guidance exists).

State how to verify readiness concretely (which endpoint or command signals success) rather than only verifying teardown by PD unreachability.

DimensionReasoningScore

Conciseness

The body is lean with zero padding: 'Always start playground in the background, verify readiness, run scoped tests, then clean up process and data' assumes Claude's competence and omits nothing needed for orientation. Every line carries load, matching the every-token-earns-its-place anchor.

5 / 5

Actionability

Concrete, actionable guidance is present: specific values and flags ('PD_ADDR=127.0.0.1:2379', '--pd.port', '--port-offset', '-run <TestName> and target subdir only'). It is not 5 because no actual start or readiness-check command appears inline — the canonical commands are deferred to docs/agents/testing-flow.md, so the body alone is not copy-paste executable.

4 / 5

Workflow Clarity

A clear sequence with explicit validation checkpoints: readiness verification before tests and teardown verification ('confirming the PD endpoint is unreachable after cleanup'). Validation steps are present so the destructive/batch cap does not apply, but there is no error-recovery feedback loop (e.g. what to do if readiness never succeeds), which the 5 anchor expects.

4 / 5

Progressive Disclosure

Well-organized sections (Overview, Workflow Checklist) with a clearly signaled, one-level-deep reference to 'docs/agents/testing-flow.md -> RealTiKV tests'. It is below 5 because the referenced detail file lives outside the skill bundle (no references/ directory exists), so the skill depends on a repo path it does not carry and cannot guarantee.

4 / 5

Total

17

/

20

Passed

Description

78%Weight 40%Scale 1-5

Based on the skill's description, can an agent find and select it at the right time? Clear, specific descriptions lead to better discovery.

A strong, niche-specific description with an explicit 'Use when' trigger and a stated lifecycle scope (startup, readiness checks, cleanup). Its main weakness is that the capability statement is embedded in the when-clause and lacks synonym-level trigger coverage.

Suggestions

Lead with a standalone capability statement (e.g. 'Starts a local TiUP playground in the background, verifies readiness, runs scoped realtikvtest tests, and tears down process and data.') before the 'Use when' clause to make the 'what' explicit.

Add trigger synonyms and variations such as 'realtikv', 'real-tikv tests', or 'integration tests' to broaden natural keyword coverage.

DimensionReasoningScore

Specificity

Names the domain ('running tests under tests/realtikvtest', 'local TiUP playground lifecycle') and several concrete actions ('strict startup, readiness checks, and cleanup'), matching the several-specific-actions-with-minor-gaps anchor. It falls short of 5 because the actions are named without any operational detail, and above 3 because it covers more than 1-2 actions.

4 / 5

Completeness

Both parts are present: an explicit 'Use when running tests under tests/realtikvtest' trigger and a stated what ('TiUP playground lifecycle with strict startup, readiness checks, and cleanup'). It is below 5 because the what is subordinate inside the when-clause rather than a clear standalone capability statement.

4 / 5

Trigger Term Quality

Good natural keyword coverage ('running tests', 'tests/realtikvtest', 'TiUP playground', 'readiness checks') that a developer in this repo would plausibly say, but common variations and synonyms are missing (e.g. 'realtikv', 'real-tikv', 'integration tests'), so it is not the comprehensive-synonym coverage of 5.

4 / 5

Distinctiveness Conflict Risk

The description occupies a clear niche (realtikvtest + TiUP playground lifecycle) with distinct, specific triggers that would not plausibly fire for any other skill, matching the minimal-conflict-risk anchor.

5 / 5

Total

17

/

20

Passed

Validation

100%

Checks the skill against the spec for correct structure and formatting. All validation checks must pass before discovery and implementation can be scored.

Validation — 16 / 16 Passed

Validation for skill structure

No warnings or errors.

Repository
pingcap/tidb
Reviewed

Table of Contents

Is this your skill?

If you maintain this skill, you can claim it as your own. Once claimed, you can manage eval scenarios, bundle related skills, attach documentation or rules, and ensure cross-agent compatibility.