CtrlK
BlogDocsLog inGet started
Tessl Logo

run-helix-tests

Submit and monitor .NET MAUI unit tests on Helix infrastructure. Supports running XAML, Resizetizer, Core, Essentials, and other unit test projects on distributed Helix queues.

64

Quality

76%

Does it follow best practices?

Run evals on this skill

Adds up to 20 points to the overall score

View guide

SecuritybySnyk

Low

Low-risk findings worth noting

Fix and improve this skill with Tessl

tessl review fix ./.github/skills/run-helix-tests/SKILL.md
SKILL.md
Quality
Evals
Security

Quality

Content

82%Weight 40%Scale 1-5

Reviews the quality of instructions and guidance provided to agents. Good implementation is clear, handles edge cases, and produces reliable results.

A strong operational skill body: fully executable, parameter-accurate commands, a clear submit-monitor-debug workflow with built-in verification, and a useful troubleshooting table. The main improvement opportunities are trimming the duplication between the Scripts and Workflow sections and moving the static project/queue tables into a reference file.

DimensionReasoningScore

Conciseness

Nearly every section is operational (commands, tables of projects/queues, troubleshooting) with no explanations of concepts Claude already knows. Minor tightening is possible: the "Scripts" section and the "Workflow" section repeat the same three invocations, and "Quick Start (Manual)" duplicates the wrapper-script path. Above anchor 3 (some unnecessary explanation) but short of anchor 5's every-token-earns-its-place.

4 / 5

Actionability

Every step is a copy-paste-ready pwsh command with real script paths and parameters that match the bundled scripts (verified: Submit-HelixTests.ps1 takes -Configuration/-Queue; Get-HelixJobStatus.ps1 takes -JobId/-Wait; Get-HelixWorkItemLog.ps1 takes -JobId/-WorkItem/-FailedOnly). Common cases are covered — submit variants, wait-for-completion, failed-only log retrieval, and a concrete debug example with a real work-item name. Matches anchor 5.

5 / 5

Workflow Clarity

The "Submit and Monitor Tests" workflow is clearly sequenced (submit → -Wait monitor → -FailedOnly log check) with verification built in via polling and failure listing, and the "Debug a Specific Test Failure" section gives a second recovery path; the "Common Issues" table adds error recovery. It is not 5 because there is no explicit checkpoint after submission (e.g., confirm the job ID was captured and the job started before waiting), and the batch submission has no stated abort/retry guidance on partial queue failures.

4 / 5

Progressive Disclosure

Well-organized sections (When to Use, Prerequisites, Scripts, Workflow, Common Issues) with the three bundle scripts referenced by exact path — all verified to exist — and the device-test topic correctly deferred to a separate instructions file. It is not 5 because the body is ~170 lines with reference-grade material (full test-project table, queue table, API URL patterns) inlined that could live in a reference file, and the device-test pointer (".github/instructions/helix-device-tests.instructions.md") is not part of the skill bundle.

4 / 5

Total

17

/

20

Passed

Description

70%Weight 40%Scale 1-5

Based on the skill's description, can an agent find and select it at the right time? Clear, specific descriptions lead to better discovery.

A specific, well-scoped description with strong natural keywords and virtually no conflict risk, but it completely lacks a "when to use" trigger clause, so an agent relying on the description alone cannot decide when to invoke the skill. Adding a sentence like "Use when the user asks to run or validate MAUI unit tests on Helix" would lift it into the top band.

Suggestions

Append an explicit trigger clause, e.g. "Use when the user asks to run, submit, or validate .NET MAUI unit tests on Helix infrastructure, or to debug Helix test failures."

Include a couple of natural user phrasings ("run tests on Helix", "test on multiple platforms") as trigger synonyms in the description.

Briefly enumerate the monitor-side actions (check job status, fetch work-item console logs, list failures) so the capability list is comprehensive rather than implied by "monitor".

DimensionReasoningScore

Specificity

Concrete actions are named ("Submit and monitor .NET MAUI unit tests on Helix infrastructure") with specific scope ("XAML, Resizetizer, Core, Essentials, and other unit test projects on distributed Helix queues") — several specific capabilities, though "monitor" is not broken into status/log actions. It sits above anchor 3 (only 1-2 generic actions) and just below anchor 5 (comprehensive action list including retrieval/diagnostics).

4 / 5

Completeness

The "what" is clear (submit and monitor unit tests on Helix), but there is no "Use when..." clause or equivalent trigger guidance anywhere in the description — the triggering guidance only exists in the body's "When to Use" section, which frontmatter-based discovery never sees. Per the judging guideline, a missing explicit trigger clause caps completeness at 3; it is not 2 because the "what" is specific, not vague.

3 / 5

Trigger Term Quality

Good natural keywords a MAUI developer would say: "MAUI", "unit tests", "Helix", project names ("XAML", "Resizetizer", "Essentials"), "queues". Missing common phrasings like "run tests", "CI", "test on Windows/macOS", and there is no explicit trigger phrasing at all. Above anchor 3 (relevant keywords but missing variations) but not the comprehensive synonym coverage of anchor 5.

4 / 5

Distinctiveness Conflict Risk

Clear niche: ".NET MAUI unit tests" + "Helix infrastructure" + named queues is highly specific and unlikely to fire for any other skill. It even distinguishes unit tests from device tests via project naming. Minimal conflict risk, matching anchor 5.

5 / 5

Total

16

/

20

Passed

Validation

100%

Checks the skill against the spec for correct structure and formatting. All validation checks must pass before discovery and implementation can be scored.

Validation — 16 / 16 Passed

Validation for skill structure

No warnings or errors.

Repository
dotnet/maui
Reviewed

Table of Contents

Is this your skill?

If you maintain this skill, you can claim it as your own. Once claimed, you can manage eval scenarios, bundle related skills, attach documentation or rules, and ensure cross-agent compatibility.