CtrlK
BlogDocsLog inGet started
Tessl Logo

cpp-testing

Use only when writing/updating/fixing C++ tests, configuring GoogleTest/CTest, diagnosing failing or flaky tests, or adding coverage/sanitizers.

63

Quality

75%

Does it follow best practices?

Run evals on this skill

Adds up to 20 points to the overall score

View guide

SecuritybySnyk

Low

Low-risk findings worth noting

Fix and improve this skill with Tessl

tessl review fix ./.kiro/skills/cpp-testing/SKILL.md
SKILL.md
Quality
Evals
Security

Quality

Content

67%Weight 40%Scale 1-5

Reviews the quality of instructions and guidance provided to agents. Good implementation is clear, handles edge cases, and produces reliable results.

A well-organized, highly actionable C++ testing reference with executable gtest/gmock/CMake/CTest examples and clearly fenced optional material. The main weakness is redundancy: the guardrails, DO/DON'T, and Common Pitfalls sections restate the same advice multiple times, and some advanced recipes could be split into reference files.

Suggestions

Consolidate 'Flaky Tests Guardrails', the DO/DON'T lists, and 'Common Pitfalls' into a single section — sleep avoidance, unique temp directories, and over-mocking each currently appear two to three times.

Move the detailed coverage and sanitizer CMake/bash recipes into a reference file (e.g., references/coverage-and-sanitizers.md), keeping only a short pointer and the enable flags inline.

Either make the UserStore fixture example fully executable against a concrete type or trim it to the fixture pattern alone, so the main examples are uniformly copy-paste ready.

DimensionReasoningScore

Conciseness

The body is mostly lean bullets and code, but the same guidance repeats across three sections: sleep avoidance appears in 'Flaky Tests Guardrails', 'DON'T', and 'Common Pitfalls'; unique temp directories and over-mocking each appear twice. This duplication is more than the minor trimming of the anchor 4 example, fitting 'mostly efficient but could be tightened'.

3 / 5

Actionability

Most examples are concrete and executable — gtest TEST/TEST_F, gmock MOCK_METHOD, the CMake/CTest quickstart, and the GCC/Clang coverage and sanitizer command chains are copy-paste ready. Two examples are labeled stubs ('Pseudocode stub: replace UserStore/User with project types', the libFuzzer harness), which is explicitly justified, but keeps it below the fully-executable anchor 5; it is well above anchor 3 since pseudocode is the exception, not the rule.

4 / 5

Workflow Clarity

The TDD loop (RED → GREEN → REFACTOR) is clearly sequenced, and the debugging workflow forms a feedback loop: 'Re-run the single failing test with gtest filter' → fix root cause → 'Expand to full suite once the root cause is fixed'. It is not 5 because validation checkpoints are implicit (no explicit 'verify the test passes' step or command) rather than explicit validation steps with error-recovery instructions.

4 / 5

Progressive Disclosure

The skill is a single well-sectioned file with no nested references — headers are clear and the fuzzing section is explicitly fenced as an 'Optional Appendix: Only use if the project already supports...'. It is not 5 because content such as the full coverage/sanitizer CMake recipes could arguably live in separate reference files for a leaner overview; it is not 3 because structure and signaling are good and nothing is buried.

4 / 5

Total

15

/

20

Passed

Description

83%Weight 40%Scale 1-5

Based on the skill's description, can an agent find and select it at the right time? Clear, specific descriptions lead to better discovery.

A strong trigger-focused description with concrete, comprehensive action enumeration and an explicit 'Use only when' clause. Its main weakness is that the capability statement ('what') is folded into the trigger list rather than stated separately, and a few natural synonyms (unit tests, gmock) are absent.

DimensionReasoningScore

Specificity

The description lists multiple concrete actions — 'writing/updating/fixing C++ tests', 'configuring GoogleTest/CTest', 'diagnosing failing or flaky tests', 'adding coverage/sanitizers' — with comprehensive coverage of the skill's domain, matching the anchor for multiple specific concrete actions. It is not 4 because the actions enumerated cover the full range of what the skill addresses rather than leaving minor gaps.

5 / 5

Completeness

The 'when' is explicit and strong ('Use only when...'), and the 'what' is conveyed through the concrete action verbs, but there is no separate capability statement of what the skill does. It is not 5 because the what is implied by the trigger list rather than clearly stated alongside the when; it is not 3 because both elements are effectively present, just merged into one clause.

4 / 5

Trigger Term Quality

Good natural keyword coverage — 'C++ tests', 'GoogleTest', 'CTest', 'flaky tests', 'coverage', 'sanitizers' — but common variations users would say such as 'unit tests', 'gmock', or file extensions are missing. It is not 5 because the anchor requires comprehensive synonyms and extensions, and not 3 because the present keywords are natural and well-targeted rather than partially relevant.

4 / 5

Distinctiveness Conflict Risk

The description carves out a clear niche — C++ testing with GoogleTest/CTest, flaky-test diagnosis, coverage and sanitizers — with trigger phrases unlikely to fire for unrelated skills. It is not 4 because it shows minimal overlap even with closely adjacent skills, satisfying the clear-niche anchor.

5 / 5

Total

18

/

20

Passed

Validation

93%

Checks the skill against the spec for correct structure and formatting. All validation checks must pass before discovery and implementation can be scored.

Validation — 15 / 16 Passed

Validation for skill structure

CriteriaDescriptionResult

frontmatter_unknown_keys

Unknown frontmatter key(s) found; consider removing or moving to metadata

Warning

Total

15

/

16

Passed

Repository
affaan-m/ECC
Reviewed

Table of Contents

Is this your skill?

If you maintain this skill, you can claim it as your own. Once claimed, you can manage eval scenarios, bundle related skills, attach documentation or rules, and ensure cross-agent compatibility.