CtrlK
BlogDocsLog inGet started
Tessl Logo

ab-testing

When the user wants to plan, design, or implement an A/B test or experiment, or build a growth experimentation program.

66

Quality

80%

Does it follow best practices?

Run evals on this skill

Adds up to 20 points to the overall score

View guide

SecuritybySnyk

Passed

No findings from the security scan

Fix and improve this skill with Tessl

tessl review fix ./plugins/AI-Agents-Safe-Coding-Skills-claude/skills/ab-testing/SKILL.md

The canonical home for this skill is ab-testing in administrakt0r/AI-Agents-Safe-Coding-Skills

SKILL.md
Quality
Evals
Security

Quality

Content

85%Weight 40%Scale 1-5

Reviews the quality of instructions and guidance provided to agents. Good implementation is clear, handles edge cases, and produces reliable results.

The body is a well-organized, actionable overview with copy-paste-ready templates, explicit checklists, and a clear experiment-loop feedback cycle. It appropriately offloads detailed sample-size and template material to two real reference files, with only minor over-explanation of statistical basics.

DimensionReasoningScore

Conciseness

Mostly lean and well-structured with tables and bullets, with only minor over-explanation of statistical basics Claude already knows (e.g., '95% confidence = p-value < 0.05' and the peeking-problem lecture).

4 / 5

Actionability

Provides concrete, usable guidance including copy-paste-ready hypothesis and playbook templates, sample-size tables, and tool lists, though as an instruction skill it lacks a fully worked end-to-end statistical example.

4 / 5

Workflow Clarity

The test lifecycle is clearly sequenced from assessment through documentation, with explicit pre-launch and analysis checklists plus a repeating experiment-loop feedback cycle.

5 / 5

Progressive Disclosure

The body is a clear overview that points via descriptive one-level-deep links to the two real reference files (sample-size-guide.md, test-templates.md) for detailed tables and templates, with content appropriately split.

5 / 5

Total

18

/

20

Passed

Description

75%Weight 40%Scale 1-5

Based on the skill's description, can an agent find and select it at the right time? Clear, specific descriptions lead to better discovery.

The description is clear, specific, and uses an explicit trigger clause covering both what the skill does and when to use it. It is concise without padding and distinct from adjacent skills, though it could add synonyms like 'split test' and separate the what/when more crisply.

DimensionReasoningScore

Specificity

Names the domain and lists several concrete actions ('plan, design, or implement an A/B test or experiment' and 'build a growth experimentation program'), but they remain process-level verbs rather than the granular multi-action coverage of the 5 anchor.

4 / 5

Completeness

Has an explicit 'When the user wants to...' trigger clause and conveys the 'what' through its action verbs, but the what is merged into the when phrasing rather than stated as a distinct standalone capability like the 5 anchor.

4 / 5

Trigger Term Quality

Includes natural user-facing terms ('A/B test', 'experiment', 'growth experimentation program'), but omits common synonyms such as 'split test' that the 5 anchor would include.

4 / 5

Distinctiveness Conflict Risk

Carves a clear experimentation niche with distinct triggers and minimal conflict risk, with only minor adjacency to related CRO/analytics skills the skill itself lists.

4 / 5

Total

16

/

20

Passed

Validation

93%

Checks the skill against the spec for correct structure and formatting. All validation checks must pass before discovery and implementation can be scored.

Validation15 / 16 Passed

Validation for skill structure

CriteriaDescriptionResult

frontmatter_unknown_keys

Unknown frontmatter key(s) found; consider removing or moving to metadata

Warning

Total

15

/

16

Passed

Repository
administrakt0r/AI-Agents-Safe-Coding-Skills
Reviewed

Table of Contents

Is this your skill?

If you maintain this skill, you can claim it as your own. Once claimed, you can manage eval scenarios, bundle related skills, attach documentation or rules, and ensure cross-agent compatibility.