CtrlK
BlogDocsLog inGet started
Tessl Logo

create-flet-control-integration-tests

Use when asked to create or update integration tests for any Flet control in sdk/python/packages/flet/integration_tests, including visual goldens and interactive behavior tests.

72

Quality

90%

Does it follow best practices?

Run evals on this skill

Adds up to 20 points to the overall score

View guide
SecuritybySnyk

Passed

No findings from the security scan

SKILL.md
Quality
Evals
Security

Quality

Content

88%Weight 40%Scale 1-5

Reviews the quality of instructions and guidance provided to agents. Good implementation is clear, handles edge cases, and produces reliable results.

A strong, execution-focused skill body: deterministic-authoring rules, complete templates for both pytest loop scopes, exact run commands, and a workflow with an explicit golden-regeneration feedback loop. The only notable weaknesses are a checklist that duplicates rules stated earlier and an unannotated references list.

DimensionReasoningScore

Conciseness

The body is mostly lean — bullet lists, two compact templates, and exact run commands, with no explanation of concepts Claude already knows. It loses a point to redundancy: the Quality checklist restates the scope-split, naming, and determinism rules already given in their own sections, and the 'When to use' section repeats the frontmatter description.

4 / 5

Actionability

Both fixture-scope templates are complete, runnable Python (imports, decorator, fixture signature, assertion), the run commands are copy-paste ready including the `-k` subset and `FLET_TEST_GOLDEN=1` golden-regeneration variants, and assertion patterns cover visual, full-page, and functional cases. Not a 4: the templates and commands cover the common cases end-to-end with nothing pseudocode about them.

5 / 5

Workflow Clarity

The 8-step authoring workflow is clearly sequenced and includes an explicit validation checkpoint ('Run target test file') plus a feedback loop for error recovery ('If expected visuals changed, regenerate goldens with FLET_TEST_GOLDEN=1 and re-run without golden mode'), backed by a Quality checklist that guards the semi-destructive golden regeneration. Not a 4: validation, retry loop, and checklist are all present, not just most.

5 / 5

Progressive Disclosure

Sections are well organized (placement, fixtures, workflow, patterns, templates, commands, checklist, references) and the reference list is one level deep pointing at real repo files. Not a 5: no bundle files exist, so all 160+ lines are inline, and the References section lists seven paths without annotating what each file provides, leaving navigation slightly under-signaled.

4 / 5

Total

18

/

20

Passed

Description

87%Weight 40%Scale 1-5

Based on the skill's description, can an agent find and select it at the right time? Clear, specific descriptions lead to better discovery.

A well-formed trigger description: explicit 'Use when' clause, concrete actions, natural terminology, and a directory path that makes it highly distinct from other skills. It could broaden trigger coverage slightly (screenshot/regression phrasing) and mention more of what the skill actually handles.

DimensionReasoningScore

Specificity

Concrete actions are named — 'create or update integration tests', 'visual goldens', 'interactive behavior tests' — anchored to a specific directory path. Not a 5: it omits capabilities the body covers (regression tests, property-behavior checks, non-visual functional assertions); not a 3: it goes well beyond naming the domain to several specific actions.

4 / 5

Completeness

Both halves are explicit in one sentence: what ('create or update integration tests... including visual goldens and interactive behavior tests') and when ('Use when asked to create or update integration tests for any Flet control in sdk/python/packages/flet/integration_tests'). The when clause names concrete trigger phrases and an exact path; the 4 anchor ('when could be more explicit') does not apply here.

5 / 5

Trigger Term Quality

Natural phrases a user would actually say are present: 'integration tests', 'create or update', 'Flet control', 'visual goldens', 'interactive behavior tests'. Not a 5: common variations users might say — 'screenshot tests', 'regression tests', 'test a control' — are absent; not a 3: coverage of the natural terminology is good, not partial.

4 / 5

Distinctiveness Conflict Risk

The description is pinned to a single niche — Flet controls and one exact integration_tests directory — with triggers ('visual goldens', 'interactive behavior tests') that no neighboring skill (unit tests, docs, general pytest) would claim. Conflict risk is minimal.

5 / 5

Total

18

/

20

Passed

Validation

100%

Checks the skill against the spec for correct structure and formatting. All validation checks must pass before discovery and implementation can be scored.

Validation — 16 / 16 Passed

Validation for skill structure

No warnings or errors.

Repository
flet-dev/flet
Reviewed

Table of Contents

Is this your skill?

If you maintain this skill, you can claim it as your own. Once claimed, you can manage eval scenarios, bundle related skills, attach documentation or rules, and ensure cross-agent compatibility.