CtrlK
BlogDocsLog inGet started
Tessl Logo

awt-e2e-testing

AI-powered E2E web testing — eyes and hands for AI coding tools. Declarative YAML scenarios, Playwright execution, visual matching (OpenCV + OCR), platform auto-detection (Flutter/React/Vue), learning DB. Install: npx skills add ksgisang/awt-skill --skill awt -g

51

Quality

56%

Does it follow best practices?

Run evals on this skill

Adds up to 20 points to the overall score

View guide

SecuritybySnyk

Passed

No findings from the security scan

Fix and improve this skill with Tessl

tessl review fix ./plugins/AI-Agents-Safe-Coding-Skills/skills/awt-e2e-testing/SKILL.md

The canonical home for this skill is awt-e2e-testing in administrakt0r/AI-Agents-Safe-Coding-Skills

SKILL.md
Quality
Evals
Security

Quality

Content

47%Weight 40%Scale 1-5

Reviews the quality of instructions and guidance provided to agents. Good implementation is clear, handles edge cases, and produces reliable results.

The content is a concise, well-structured overview but functions more as a marketing README than actionable skill guidance — it lacks example scenarios, run commands, and a defined workflow with validation.

Suggestions

Add a minimal runnable example: a short YAML scenario and the command to execute it via AWT.

Provide a sequenced workflow (design scenario → run → read failure diagnosis → consult learning DB) with an explicit validation/checkpoint step.

Trim the closing marketing lines ('Built with the help of AI coding tools...', 'Actively developed by a solo developer...') to recover tokens.

DimensionReasoningScore

Conciseness

The body is lean and assumes Claude's competence ('YAML scenarios → Playwright with human-like interaction'); only minor marketing fluff ('Built with the help of AI coding tools...', 'Feedback welcome!') could be trimmed, fitting 4 rather than the perfectly-lean 5.

4 / 5

Actionability

Beyond the install command there is no executable guidance — no example YAML scenario, no run command, no API — only high-level hints about what works, matching the 'minimal concrete guidance; missing the specific steps' anchor at 2.

2 / 5

Workflow Clarity

No sequenced workflow is given; 'Your AI designs YAML test scenarios; AWT executes them with Playwright' is a rough two-beat sketch with steps poorly defined and no validation checkpoints, fitting 2.

2 / 5

Progressive Disclosure

The body is short (~25 lines) and well-organized into 'What works now' and 'Links' sections, but references point to external GitHub repos rather than one-level-deep bundle files, so good-but-not-ideal structure fits 4.

4 / 5

Total

12

/

20

Passed

Description

66%Weight 40%Scale 1-5

Based on the skill's description, can an agent find and select it at the right time? Clear, specific descriptions lead to better discovery.

The description is specific and reasonably distinctive, clearly stating what the skill does, but it lacks an explicit 'Use when...' trigger clause, which caps its completeness and limits natural-term coverage.

Suggestions

Add an explicit 'Use when...' trigger clause (e.g., 'Use when the user needs end-to-end or browser testing of a web app, visual UI matching, or AI-driven test scenarios').

Include more natural user-facing trigger phrases such as 'end-to-end tests', 'browser testing', and 'test my web app' alongside the technical keywords.

Drop or trim the install command from the description to keep it focused on capabilities and triggers.

DimensionReasoningScore

Specificity

Lists several concrete capabilities ('Declarative YAML scenarios, Playwright execution, visual matching (OpenCV + OCR), platform auto-detection (Flutter/React/Vue), learning DB') — multiple specific actions with only minor coverage gaps, fitting the 4 anchor rather than the comprehensive 5.

4 / 5

Completeness

The 'what' is clear (AI-powered E2E web testing with enumerated features) but there is no 'Use when...' clause or equivalent trigger guidance, and the rubric caps completeness at 3 when that is missing.

3 / 5

Trigger Term Quality

Includes natural terms like 'E2E web testing' alongside technical keywords (Playwright, OpenCV, OCR), but misses common natural phrasings such as 'end-to-end tests', 'browser testing', or 'test my app', so good-but-not-comprehensive coverage fits 4.

4 / 5

Distinctiveness Conflict Risk

The niche ('eyes and hands for AI coding tools' / E2E web testing via YAML + Playwright) is mostly distinct with only minor overlap risk against other testing skills, fitting 4 rather than the fully-distinct 5.

4 / 5

Total

15

/

20

Passed

Validation

93%

Checks the skill against the spec for correct structure and formatting. All validation checks must pass before discovery and implementation can be scored.

Validation15 / 16 Passed

Validation for skill structure

CriteriaDescriptionResult

frontmatter_unknown_keys

Unknown frontmatter key(s) found; consider removing or moving to metadata

Warning

Total

15

/

16

Passed

Repository
administrakt0r/AI-Agents-Safe-Coding-Skills
Reviewed

Table of Contents

Is this your skill?

If you maintain this skill, you can claim it as your own. Once claimed, you can manage eval scenarios, bundle related skills, attach documentation or rules, and ensure cross-agent compatibility.