CtrlK
BlogDocsLog inGet started
Tessl Logo

daytona-flow-validator

do e2e tests, validate feature, prove it works, pass/fail, frame proof, screenshots, CDP assertions. Daytona validation loop for real app behavior with repair before declaring success.

65

Quality

77%

Does it follow best practices?

Run evals on this skill

Adds up to 20 points to the overall score

View guide

SecuritybySnyk

Medium

Suggest reviewing before use

Fix and improve this skill with Tessl

tessl review fix ./.opencode/skills/daytona-flow-validator/SKILL.md
SKILL.md
Quality
Evals
Security

Quality

Content

85%Weight 40%Scale 1-5

Reviews the quality of instructions and guidance provided to agents. Good implementation is clear, handles edge cases, and produces reliable results.

The body is highly actionable and exceptionally strong on workflow clarity with validation checkpoints and repair feedback loops. Its main weakness is progressive disclosure: a long monolithic file with no reference files for the detailed bash/JS recipes.

Suggestions

Extract the detailed recipe blocks (Lexical Composer JS, Linux Desktop Automation xdotool/wmctrl patterns) into reference files under references/ and link to them one level deep from the main body.

Consolidate the repeated native-dialog/Authorize-folder guardrail into a single canonical check referenced from each section that needs it.

Consider moving the per-section screenshot visual-checks list into a single checklist reference to reduce inline length while keeping the rule visible.

DimensionReasoningScore

Conciseness

Tight, domain-specific guidance with no padding of concepts Claude already knows, though the native-dialog/picker guardrail recurs across the Linux Desktop Automation, Screenshots, and Repair Loop sections and could be consolidated.

4 / 5

Actionability

Provides copy-paste-ready, properly-escaped code and commands (Lexical paste JS, xdotool/wmctrl bash, browser_screenshot calls) covering the common cases with clearly marked placeholders.

5 / 5

Workflow Clarity

Explicitly sequenced observe-act-observe-assert loop, a 5-step Repair Loop feedback path, checklists for pass evidence and screenshot visual checks, and a defined Final Verdict scale.

5 / 5

Progressive Disclosure

Well-organized with clear section headers, but it is a ~245-line monolithic file with detailed recipes inlined and no one-level-deep reference files, so content that could be split remains inline.

3 / 5

Total

17

/

20

Passed

Description

70%Weight 40%Scale 1-5

Based on the skill's description, can an agent find and select it at the right time? Clear, specific descriptions lead to better discovery.

The description is specific and clearly niche-scoped to Daytona flow validation, but it omits any explicit 'Use when...' trigger guidance, which caps its completeness. Trigger terms are mostly natural with some jargon.

Suggestions

Add an explicit 'Use when...' clause (e.g., 'Use when validating a Daytona Electron or browser flow end-to-end, or when the user asks to prove a feature works on Daytona').

Soften jargon-heavy terms ('CDP assertions', 'frame proof') with user-natural synonyms, or pair each with a plain-language phrase.

Lead with the core action verb phrase in third person before the keyword list so the 'what' reads as a sentence rather than a comma list.

DimensionReasoningScore

Specificity

Lists several concrete actions ('do e2e tests', 'frame proof', 'screenshots', 'CDP assertions') but they cluster within the validation/proof subdomain rather than spanning distinct capabilities.

4 / 5

Completeness

The 'what' is clear and detailed, but there is no explicit 'Use when...' clause or equivalent trigger guidance, which caps completeness at 3 per the rubric guideline.

3 / 5

Trigger Term Quality

Includes natural phrases a user would say ('e2e tests', 'prove it works', 'pass/fail', 'screenshots') alongside some jargon ('CDP assertions', 'frame proof'), leaving a few common synonyms uncovered.

4 / 5

Distinctiveness Conflict Risk

Names a clear niche (Daytona, CDP, Electron, browser-flow validation) with distinct triggers and minimal overlap risk with other skills.

5 / 5

Total

16

/

20

Passed

Validation

100%

Checks the skill against the spec for correct structure and formatting. All validation checks must pass before discovery and implementation can be scored.

Validation16 / 16 Passed

Validation for skill structure

No warnings or errors.

Repository
Devin-AXIS/iPolloWork
Reviewed

Table of Contents

Is this your skill?

If you maintain this skill, you can claim it as your own. Once claimed, you can manage eval scenarios, bundle related skills, attach documentation or rules, and ensure cross-agent compatibility.