CtrlK
BlogDocsLog inGet started
Tessl Logo

ab-test-store-listing

When the user wants to A/B test App Store product page elements to improve conversion rate. Also use when the user mentions "A/B test", "product page optimization", "test my screenshots", "test my icon", "conversion rate optimization", "CPP", or "custom product pages". For screenshot design, see screenshot-optimization. For metadata optimization, see metadata-optimization.

64

Quality

76%

Does it follow best practices?

Run evals on this skill

Adds up to 20 points to the overall score

View guide

SecuritybySnyk

Passed

No findings from the security scan

Fix and improve this skill with Tessl

tessl review fix ./skills/ab-test-store-listing/SKILL.md
SKILL.md
Quality
Evals
Security

Quality

Content

75%Weight 40%Scale 1-5

Reviews the quality of instructions and guidance provided to agents. Good implementation is clear, handles edge cases, and produces reliable results.

A well-structured, domain-rich body: nearly everything is non-obvious App Store-specific data delivered via tables, with a clear five-step testing workflow, prioritization matrix, and concrete output templates. Weaknesses are minor — a placeholder sample-size calculation with no real formula, slight redundancy between the framework and test-idea tables, and no explicit loop for handling inconclusive tests.

Suggestions

Replace the fill-in sample-size template with an actual rule-of-thumb formula or a small worked example (e.g., 'required impressions per variant ≈ 16 × p × (1-p) / MDE² for 95% confidence') so Step 3 is executable rather than a placeholder.

Add a short 'If the test is inconclusive' branch to Step 5 (e.g., extend duration once, then fall back to the control and move to the next priority test) to close the error-recovery gap in the workflow.

Trim the 'You are an expert...' role-play opener and merge overlapping guidance between the Test Design Framework and the Common Test Ideas tables to tighten token efficiency.

DimensionReasoningScore

Conciseness

The body is dominated by dense tables and terse bullets carrying non-obvious domain facts (PPO limits, '35 custom product pages', '90% confidence minimum', '80% of users never scroll past the first 3 screenshots'), which is exactly what Claude doesn't already know. Minor trimmable padding remains — the role-play opener 'You are an expert in App Store product page optimization...' and partial overlap between the Test Design Framework and the Common Test Ideas tables — so it fits anchor 4 ('efficient; minor instances of over-explanation that could be trimmed') rather than anchor 5.

4 / 5

Actionability

Guidance is concrete for an instruction-only skill: exact hypothesis template ('If we [change], then [metric] will [improve] because [reason]'), the App Store Connect click path ('Go to Product Page Optimization → Create a new test → Upload variant assets'), and a copy-paste Test Plan output template. It falls short of anchor 5 because the sample-size step is a fill-in placeholder ('Required sample per variant: ~[N] impressions') with no actual formula or worked example, leaving the key sizing step pseudocode-like.

4 / 5

Workflow Clarity

The five-step Test Design Framework (Hypothesis → Variants → Sample Size → Run → Interpret) is clearly sequenced, with decision checkpoints like 'Aim for 95% confidence before making decisions', 'Monitor but don't stop early', 'Change ONE thing per test', and a results-interpretation sequence that feeds the next test and a 3-month roadmap. Not anchor 5: there is no explicit error-recovery loop (e.g., what to do when a test is inconclusive or significance is never reached), a minor validation gap characteristic of anchor 4.

4 / 5

Progressive Disclosure

No bundle files exist (no references/, scripts/, or assets/ directories), so the skill is a single well-organized file with clear one-level section headers and a Related Skills section pointing to sibling skills (screenshot-optimization, metadata-optimization, app-analytics, aso-audit). The one external reference (app-marketing-context.md) is clearly signaled in the Initial Assessment. It is not a 5 because at ~220 lines the Common Test Ideas tables and expected-impact data are bulk detail that could arguably live in a one-level-deep reference file, a minor organization gap matching anchor 4.

4 / 5

Total

16

/

20

Passed

Description

78%Weight 40%Scale 1-5

Based on the skill's description, can an agent find and select it at the right time? Clear, specific descriptions lead to better discovery.

A strong description with an explicit 'when the user wants...' clause, a rich list of natural trigger terms including jargon like 'CPP' and colloquial 'test my icon', and clear boundary pointers to sibling skills. The main gap is that the 'what' clause names only one capability rather than enumerating the skill's concrete actions (test design, sample-size estimation, results interpretation, roadmapping).

DimensionReasoningScore

Specificity

The description names the domain precisely ("A/B test App Store product page elements to improve conversion rate") but lists only that single action — it never states the concrete things the skill does (design tests, calculate sample size, interpret results, build a testing roadmap). This matches anchor 3 ('names domain and 1-2 concrete actions, but not comprehensive') and falls short of anchor 4, which expects several specific actions enumerated.

3 / 5

Completeness

It explicitly answers 'when' twice — "When the user wants to A/B test App Store product page elements..." and "Also use when the user mentions 'A/B test', 'product page optimization', ..." — with concrete trigger phrases, and answers 'what' with a clear capability statement (A/B testing App Store product page elements to improve conversion rate). This matches anchor 5 ('clearly and explicitly answers both what AND when with concrete trigger phrases'); it is not a 4 because the 'when' is fully explicit and specific, not merely present.

5 / 5

Trigger Term Quality

Quotes like "A/B test", "product page optimization", "test my screenshots", "test my icon", "conversion rate optimization", "CPP", and "custom product pages" give good natural-phrase coverage, including the colloquial 'test my screenshots/icon' forms and domain jargon (CPP). Not a 5: common synonyms such as 'ASO', 'increase downloads', 'app store conversion', or 'icon test' are absent, so a few natural terms users might say are missing.

4 / 5

Distinctiveness Conflict Risk

The niche is clear (Apple App Store PPO/CPP testing) and the description actively disambiguates from sibling skills ("For screenshot design, see screenshot-optimization. For metadata optimization, see metadata-optimization"), reducing conflict risk. However, generic triggers like "A/B test" and "conversion rate optimization" could also fire for web-landing-page or general CRO skills, leaving minor overlap risk with closely related skills — anchor 4 rather than 5.

4 / 5

Total

16

/

20

Passed

Validation

100%

Checks the skill against the spec for correct structure and formatting. All validation checks must pass before discovery and implementation can be scored.

Validation — 16 / 16 Passed

Validation for skill structure

No warnings or errors.

Repository
Eronred/aso-skills
Reviewed

Table of Contents

Is this your skill?

If you maintain this skill, you can claim it as your own. Once claimed, you can manage eval scenarios, bundle related skills, attach documentation or rules, and ensure cross-agent compatibility.