CtrlK
BlogDocsLog inGet started
Tessl Logo

ci-pipeline-monitor

Monitors .NET runtime CI test pipelines on Azure DevOps. Use this skill when asked to monitor CI pipeline test results, triage CI test failures across ADO pipelines, or generate CI test monitoring reports.

68

Quality

84%

Does it follow best practices?

Run evals on this skill

Adds up to 20 points to the overall score

View guide

SecuritybySnyk

—

The risk profile of this skill

SKILL.md
Quality
Evals
Security

Quality

Content

81%Weight 40%Scale 1-5

Reviews the quality of instructions and guidance provided to agents. Good implementation is clear, handles edge cases, and produces reliable results.

A highly actionable, rigorously sequenced workflow: every scripted step has an exact command, validation with a bounded retry loop guards the batch DB operations, and banned/allowed tool tables prevent known failure modes. The main weakness is token efficiency — the same critical invariants are restated three times across the Scripts table, pipeline overview, and Workflow sections — and some detail-heavy content (SQL schema, retry logic) stays inlined rather than in the clearly-signaled reference files.

Suggestions

State each script's behavior once (keep the most detail in the Workflow Step sections) and trim the Scripts table plus the end-to-end pipeline comments to one-line summaries, cutting the triple restatement of extract_failed_tests.py behavior.

Move the full SQL CREATE TABLE schema and the Step 5a retry playbook into reference files (e.g., references/schema.md, references/validation-checks.md) and keep only the table names and key columns inline in SKILL.md.

Consolidate the duplicated DB-lifecycle prose (created/populated/validated/read by which step) that appears in both the Database Schema intro and the step descriptions into one place.

DimensionReasoningScore

Conciseness

The body is dense and operational — it never explains concepts Claude already knows — but it could be tightened: the Step 3 extraction behavior ('error_message, stack_trace from the ADO API', '.WorkItemExecution' stripping, generic-message handling) is restated nearly verbatim in the Scripts table, the end-to-end pipeline block, and the Workflow Step 3 section, and the SQL schema comments duplicate surrounding prose. This matches the anchor 'mostly efficient but includes some unnecessary explanation or could be tightened' better than anchor 4's 'minor instances'.

3 / 5

Actionability

Fully executable throughout: copy-paste-ready commands for every scripted step, a concrete AzDO API endpoint with api-version, an explicit DB schema, per-step allowed-tools tables, and exact WARN log formats. Not below 5 — nothing is pseudocode or abstract; even edge behavior (203 without auth, 60-minute token validity) is specified.

5 / 5

Workflow Clarity

Steps 0–7 are clearly sequenced with an explicit validation checkpoint (Step 5, exit 1 on failure) and a full feedback loop (Step 5a: read validator output, fix, re-validate, up to 3 retries, stop when the failure count stops decreasing, log remaining WARNs and proceed). This is precisely the anchor-5 'validate → fix → re-validate → only when valid proceed' pattern, and the destructive/batch cap does not apply since validation is present.

5 / 5

Progressive Disclosure

Good structure with one-level-deep, well-signaled references — each link (pipelines.md, references/prerequisites.md, references/triage-workflow.md, references/validation-checks.md, references/verbatim-rules.md, log-template.md, report-template.md) states what it contains and when to use it. Not a 5 because sizable detail is inlined in SKILL.md that arguably belongs in those references (the ~70-line SQL schema, the full Step 1 definition-resolution procedure, and the Step 5a retry playbook), which keeps it at 'minor organization gaps' per anchor 4. No bundle reference/ or scripts/ files were provided to verify the referenced paths against.

4 / 5

Total

17

/

20

Passed

Description

87%Weight 40%Scale 1-5

Based on the skill's description, can an agent find and select it at the right time? Clear, specific descriptions lead to better discovery.

A well-crafted description: third-person 'what' statement followed by an explicit 'Use when...' clause with concrete, natural trigger phrases. Domain scoping is tight and conflict risk is minimal; only minor keyword-synonym and action-coverage gaps keep specificity and trigger quality at 4.

DimensionReasoningScore

Specificity

Names the domain ('.NET runtime CI test pipelines on Azure DevOps') and several concrete actions — 'monitor CI pipeline test results', 'triage CI test failures across ADO pipelines', 'generate CI test monitoring reports'. Not a 5 because 'monitors' is somewhat generic and coverage of the skill's full capability set (e.g., regression bisecting, report generation is only weakly differentiated) has minor gaps.

4 / 5

Completeness

Clearly answers both: 'what' — 'Monitors .NET runtime CI test pipelines on Azure DevOps' — and 'when' — 'Use this skill when asked to monitor CI pipeline test results, triage CI test failures across ADO pipelines, or generate CI test monitoring reports.' Both are explicit with concrete trigger phrases, matching the anchor-5 example structure.

5 / 5

Trigger Term Quality

Includes natural phrases users would actually say: 'monitor CI pipeline test results', 'triage CI test failures', 'generate CI test monitoring reports', plus both 'Azure DevOps' and 'ADO'. Not a 5 because common variations like 'check the build', 'CI failures', or 'pipeline health' are missing.

4 / 5

Distinctiveness Conflict Risk

A clear niche (.NET runtime CI on Azure DevOps) with distinct, unambiguous triggers; unlikely to fire for unrelated CI/reporting skills. Anti-drift check: nothing pushes it below 5 — no generic wording that would overlap other skills.

5 / 5

Total

18

/

20

Passed

Validation

87%

Checks the skill against the spec for correct structure and formatting. All validation checks must pass before discovery and implementation can be scored.

Validation — 14 / 16 Passed

Validation for skill structure

CriteriaDescriptionResult

relative_links

Relative link issues: 9 missing

Warning

referenced_paths_exist

Referenced path issues: 34 missing

Warning

Total

14

/

16

Passed

Repository
dotnet/runtime
Reviewed

Table of Contents

Is this your skill?

If you maintain this skill, you can claim it as your own. Once claimed, you can manage eval scenarios, bundle related skills, attach documentation or rules, and ensure cross-agent compatibility.