CtrlK
BlogDocsLog inGet started
Tessl Logo

autolab-reporter

Operate the local Trackio reporter for Autolab HF Jobs. Use when a reporter or planner needs to inspect scores, active jobs, worker anomalies, duplicate launches, or the overall experiment board.

68

Quality

86%

Does it follow best practices?

Run evals on this skill

Adds up to 20 points to the overall score

View guide

SecuritybySnyk

Low

Low-risk findings worth noting

SKILL.md
Quality
Evals
Security

Quality

Content

93%Weight 40%Scale 1-5

Reviews the quality of instructions and guidance provided to agents. Good implementation is clear, handles edge cases, and produces reliable results.

The body is exemplary for a simple inspection skill: lean, fully actionable commands, and well-structured sections with no padding. The only gap is the absence of explicit validation/verification steps in the monitoring workflow.

Suggestions

Add a verification checkpoint after sync, e.g. confirm the job count matches expectations before relying on the summary or dashboard.

Note what "anomalies" look like concretely (e.g., a single line on expected vs. observed job states) so the "fix those before launching more work" guardrail is actionable.

Clarify when to use `--watch --interval 300` versus the one-shot sync so the continuous-reporting step has an explicit stopping/escalation condition.

DimensionReasoningScore

Conciseness

At roughly 30 lines it is lean and efficient, assumes Claude's competence, and never explains what Trackio or HF Jobs are; every section earns its place.

5 / 5

Actionability

Each workflow step is a fully executable, copy-paste-ready `uv run scripts/trackio_reporter.py` command with concrete flags, and "What To Watch" gives specific detection targets.

5 / 5

Workflow Clarity

A clear five-step sequence with concrete commands, but it lacks explicit validation checkpoints (e.g., confirming a sync succeeded or that the dashboard reflects expected jobs), keeping it below a 5.

4 / 5

Progressive Disclosure

A short, well-organized skill with clear Workflow / What To Watch / Guardrails sections and no need for external bundle files; content is appropriately self-contained and easy to navigate.

5 / 5

Total

19

/

20

Passed

Description

78%Weight 40%Scale 1-5

Based on the skill's description, can an agent find and select it at the right time? Clear, specific descriptions lead to better discovery.

The description is specific, well-scoped, and clearly distinguishes the skill's niche with explicit trigger guidance. It could push higher by broadening trigger-term synonyms and sharpening the "operate" verb into distinct operations.

Suggestions

Replace the generic verb "Operate" with the reporter's concrete operations (e.g., "Sync, summarize, and monitor Autolab HF Jobs via the local Trackio reporter").

Add common synonyms a user might say to the trigger list, such as "job queue", "leaderboard", or "parallel runs".

Split the single "Use when" clause into a short bulleted trigger list so each situation is individually scannable.

DimensionReasoningScore

Specificity

"Operate the local Trackio reporter" names the domain and "inspect scores, active jobs, worker anomalies, duplicate launches, or the overall experiment board" lists several concrete actions, though they are mostly inspect-variants rather than the reporter's full operation set.

4 / 5

Completeness

Provides a clear "what" (operate the Trackio reporter) and an explicit "Use when..." "when" clause with concrete triggers, but the "what" verb ("operate") is slightly generic and the trigger list is a single clause.

4 / 5

Trigger Term Quality

Natural fleet-management terms ("scores", "active jobs", "worker anomalies", "duplicate launches", "experiment board") give good keyword coverage, but some common user phrasings like "job queue" or "leaderboard" are absent.

4 / 5

Distinctiveness Conflict Risk

The niche ("Trackio reporter for Autolab HF Jobs") is highly specific with distinct triggers and minimal overlap risk with other skills.

5 / 5

Total

17

/

20

Passed

Validation

75%

Checks the skill against the spec for correct structure and formatting. All validation checks must pass before discovery and implementation can be scored.

Validation — 12 / 16 Passed

Validation for skill structure

CriteriaDescriptionResult

metadata_version

'metadata.version' is missing

Warning

metadata_field

'metadata' should map string keys to string values

Warning

frontmatter_unknown_keys

Unknown frontmatter key(s) found; consider removing or moving to metadata

Warning

referenced_paths_exist

Referenced path issues: 4 missing

Warning

Total

12

/

16

Passed

Repository
huggingface/context-course
Reviewed

Table of Contents

Is this your skill?

If you maintain this skill, you can claim it as your own. Once claimed, you can manage eval scenarios, bundle related skills, attach documentation or rules, and ensure cross-agent compatibility.