CtrlK
BlogDocsLog inGet started
Tessl Logo

cpp-testing

Use only when writing/updating/fixing C++ tests, configuring GoogleTest/CTest, diagnosing failing or flaky tests, or adding coverage/sanitizers.

67

Quality

81%

Does it follow best practices?

Run evals on this skill

Adds up to 20 points to the overall score

View guide

SecuritybySnyk

Low

Low-risk findings worth noting

SKILL.md
Quality
Evals
Security

Quality

Content

71%Weight 40%Scale 1-5

Reviews the quality of instructions and guidance provided to agents. Good implementation is clear, handles edge cases, and produces reliable results.

The content is highly actionable with executable, well-chosen examples and a clear workflow, but it spends tokens on redundant flakiness guidance repeated across three sections and inlines advanced reference-grade material (coverage, sanitizers, fuzzing) in the single file. Splitting reference material out and deduplicating the best-practice sections would tighten it considerably.

Suggestions

Consolidate the overlapping flakiness guidance from 'Flaky Tests Guardrails', 'Best Practices → DON'T', and 'Common Pitfalls' (sleeps, temp dirs, time/network, over-mocking each appear multiple times) into a single section.

Move the full coverage recipes, sanitizer CMake, and the fuzzing appendix into reference files (e.g. references/coverage.md, references/sanitizers.md, references/fuzzing.md) with one-line pointers from SKILL.md.

Trim 'Core Concepts' entries that restate knowledge Claude already has (TDD loop, mocks vs fakes) down to project-specific conventions only.

DimensionReasoningScore

Conciseness

The body is mostly lean code, but it re-explains concepts Claude already knows ('TDD loop: red → green → refactor', 'Mocks vs fakes') and repeats the same flakiness guidance across three sections — 'Never use `sleep` for synchronization', 'Don't use sleeps as synchronization when a condition variable can be used', and '**Flaky concurrency tests** → Use condition variables/latches' — with over-mocking also appearing three times.

3 / 5

Actionability

Nearly all examples are copy-paste executable: complete gtest/gmock test files, a full CMake quickstart with FetchContent and gtest_discover_tests(), and ready-to-run ctest, lcov, and llvm-cov command sequences. The two pseudocode blocks are explicitly labeled and justified as project-type placeholders.

5 / 5

Workflow Clarity

The RED → GREEN → REFACTOR loop and the numbered debugging sequence ending in 'Expand to full suite once the root cause is fixed' give a clear order with a closing checkpoint, but validation checkpoints are implicit rather than explicit validate-fix-retry steps.

4 / 5

Progressive Disclosure

The body is ~320 lines in a single file with no bundle or reference files, so advanced material (full GCC/Clang coverage recipes, sanitizer CMake, the fuzzing appendix) is inlined where a leaner overview pointing to reference files would fit better; section headers are otherwise clear.

3 / 5

Total

15

/

20

Passed

Description

92%Weight 40%Scale 1-5

Based on the skill's description, can an agent find and select it at the right time? Clear, specific descriptions lead to better discovery.

A strong, highly targeted description: it explicitly scopes the skill with a concrete when-clause and enumerates specific capabilities covering the C++ testing domain. The only gap is modest synonym coverage (gmock, unit tests, specific sanitizer names).

DimensionReasoningScore

Specificity

The description lists multiple concrete actions — 'writing/updating/fixing C++ tests', 'configuring GoogleTest/CTest', 'diagnosing failing or flaky tests', 'adding coverage/sanitizers' — giving comprehensive coverage of the C++ testing domain with no meaningful gaps.

5 / 5

Completeness

The 'Use only when...' clause explicitly answers when with concrete trigger situations, and the enumerated verb phrases (writing, configuring, diagnosing, adding) clearly state what the skill does.

5 / 5

Trigger Term Quality

Natural terms like 'C++ tests', 'GoogleTest', 'CTest', 'flaky tests', 'coverage', and 'sanitizers' are present and would be said by users, but common variations such as 'gmock', 'unit tests', or tool names like ASan are missing.

4 / 5

Distinctiveness Conflict Risk

The combination of 'C++ tests' with 'GoogleTest/CTest', 'flaky tests', and 'sanitizers' carves out a clear niche that is unlikely to trigger for general C++ work or other-language testing skills.

5 / 5

Total

19

/

20

Passed

Validation

93%

Checks the skill against the spec for correct structure and formatting. All validation checks must pass before discovery and implementation can be scored.

Validation — 15 / 16 Passed

Validation for skill structure

CriteriaDescriptionResult

metadata_version

'metadata.version' is missing

Warning

Total

15

/

16

Passed

Repository
affaan-m/ECC
Reviewed

Table of Contents

Is this your skill?

If you maintain this skill, you can claim it as your own. Once claimed, you can manage eval scenarios, bundle related skills, attach documentation or rules, and ensure cross-agent compatibility.