CtrlK
BlogDocsLog inGet started
Tessl Logo

develop-web-game

Use when Codex is building or iterating on a web game (HTML/JS) and needs a reliable development + testing loop: implement small changes, run a Playwright-based test script with short input bursts and intentional pauses, inspect screenshots/text, and review console errors with render_game_to_text.

68

Quality

82%

Does it follow best practices?

Run evals on this skill

Adds up to 20 points to the overall score

View guide

SecuritybySnyk

Low

Low-risk findings worth noting

SKILL.md
Quality
Evals
Security

Quality

Content

81%Weight 40%Scale 1-5

Reviews the quality of instructions and guidance provided to agents. Good implementation is clear, handles edge cases, and produces reliable results.

A highly actionable, well-sequenced skill body: every command and code pattern is executable and consistent with the bundled script and payload reference, and the workflow includes genuine validation checkpoints and feedback loops. The main cost is token efficiency — several topics are stated twice across the Workflow and dedicated sections, and one long paragraph in "Test Artifacts to Review" could be tightened.

Suggestions

Consolidate the duplicated progress.md guidance (Workflow step 5 vs. the "Progress Tracking" section) and the duplicated Playwright availability guidance (step 6 vs. "Playwright Prerequisites") into single sections referenced from the workflow.

Break the run-on "Test Artifacts to Review" paragraph into a short bulleted checklist and trim the repeated "fix and rerun in a loop until correct" phrasing to one canonical statement.

Consider moving the render_game_to_text payload guidance and Core Game Guidelines details into a reference file, keeping SKILL.md as a lean overview of the implement-act-pause-observe-adjust loop.

DimensionReasoningScore

Conciseness

Mostly efficient directive prose with no tutoring on concepts Claude already knows, but it is noticeably duplicated: progress.md handling appears in Workflow step 5 and again in "Progress Tracking", Playwright availability in step 6 and again in "Playwright Prerequisites", and "fix and rerun in a loop" phrasing recurs across four sections. Not 4 because the duplication and the long run-on "Test Artifacts to Review" paragraph could be meaningfully tightened; not 2 because there is no padded conceptual explanation.

3 / 5

Actionability

Fully executable guidance: a copy-paste node command with real flags (verified against the client script's arg parser — --url, --actions-file, --click-selector, --iterations, --pause-ms all exist), an inline actions JSON example matching references/action_payloads.json, and complete minimal patterns for both window.render_game_to_text and window.advanceTime. Not 4 because even edge details (required action-burst flags, npx check command) are documented and consistent with the shipped script.

5 / 5

Workflow Clarity

The 14-step Workflow is clearly sequenced with explicit validation checkpoints — "Review console errors and fix the first new issue before continuing", "Open the latest screenshot, verify expected visuals, fix any issues, and rerun", "Reset between scenarios" — plus feedback loops (repeat steps 7-13) and a Test Checklist. Not 4 because checkpoints, error-recovery loops, and a checklist are all explicitly present rather than merely most.

5 / 5

Progressive Disclosure

Good structure against the actual bundle: scripts/web_game_playwright_client.js and references/action_payloads.json are both real, referenced from dedicated "Scripts" and "References" sections, one level deep and clearly signaled from the Workflow. Not 5 because material duplicated between the Workflow and later sections (progress.md, Playwright prerequisites) and the inline game-guideline/pattern content could be split into references for a leaner overview; not 3 because navigation is clear and the split that exists is appropriate.

4 / 5

Total

17

/

20

Passed

Description

83%Weight 40%Scale 1-5

Based on the skill's description, can an agent find and select it at the right time? Clear, specific descriptions lead to better discovery.

A strong description: explicit 'Use when' trigger tied to web game development in HTML/JS, followed by four concrete capability statements in third-person imperative style. The main gaps are missing common synonym variants (browser/HTML5/canvas game) and omission of a couple of the skill's supporting capabilities.

DimensionReasoningScore

Specificity

Quotes several concrete actions — "implement small changes", "run a Playwright-based test script with short input bursts and intentional pauses", "inspect screenshots/text", "review console errors with render_game_to_text" — comparable to the anchor listing several specific actions with minor gaps. Not 5 because capabilities like the time-stepping hook, progress tracking, and payload references are omitted; not 3 because it lists more than 1-2 concrete actions with clear technical specifics.

4 / 5

Completeness

Explicitly answers both: what it does (implement small changes, run the Playwright test loop, inspect screenshots/text, review console errors) and when to use it ("Use when Codex is building or iterating on a web game (HTML/JS) and needs a reliable development + testing loop"). This matches the anchor with concrete trigger phrases for both what and when; no cap applies since the 'Use when' clause is present and specific.

5 / 5

Trigger Term Quality

Includes the natural phrases a user would say — "building or iterating on a web game", "HTML/JS", "test script", "screenshots" — giving good keyword coverage. Not 5 because common synonyms and variants are missing ("browser game", "HTML5 game", "canvas game", "JavaScript game"); not 3 because the primary natural terms are present rather than only generic keywords.

4 / 5

Distinctiveness Conflict Risk

The niche — a Playwright-driven development/testing loop specifically for HTML/JS web games — is mostly distinct with clearly targeted triggers. Not 5 because "building... (HTML/JS)" has minor overlap risk with general web-development or Playwright-testing skills; not 3 because the game-specific, render_game_to_text-anchored framing makes confusing it with generic skills unlikely.

4 / 5

Total

17

/

20

Passed

Validation

93%

Checks the skill against the spec for correct structure and formatting. All validation checks must pass before discovery and implementation can be scored.

Validation — 15 / 16 Passed

Validation for skill structure

CriteriaDescriptionResult

frontmatter_unknown_keys

Unknown frontmatter key(s) found; consider removing or moving to metadata

Warning

Total

15

/

16

Passed

Repository
netease-youdao/LobsterAI
Reviewed

Table of Contents

Is this your skill?

If you maintain this skill, you can claim it as your own. Once claimed, you can manage eval scenarios, bundle related skills, attach documentation or rules, and ensure cross-agent compatibility.