CtrlK
BlogDocsLog inGet started
Tessl Logo

observability-logging

Logs sessions, tracks activity, records delegation decisions, and stores review/dispute outcomes as NDJSON audit trails. Use when logging session activity, tracking work, recording decisions, building audit trails, capturing delegation history, or running pre-response verification checklists. Trigger terms: log, track activity, audit trail, session record, delegation log, NDJSON

72

Quality

88%

Does it follow best practices?

Run evals on this skill

Adds up to 20 points to the overall score

View guide

SecuritybySnyk

Passed

No findings from the security scan

SKILL.md
Quality
Evals
Security

Quality

Content

88%Weight 40%Scale 1-5

Reviews the quality of instructions and guidance provided to agents. Good implementation is clear, handles edge cases, and produces reliable results.

A highly actionable, well-structured body: five complete CLI examples with a verify step and a pre-response checklist give it strong workflow clarity and executability. The only deductions are mild triplicated gate emphasis and a schema reference that points outside the skill bundle rather than to a bundled reference file.

DimensionReasoningScore

Conciseness

The body is dense and functional — no concept explanations Claude already knows, and every section carries operational weight (event table, five concrete CLI invocations, verify command, checklist, output contract). It falls short of anchor 5 only because the logging requirement is asserted three times (the top "HARD GATE" banner, the "STOP. Verify before responding" banner, and the checklist), which is emphasis that could be trimmed — fitting anchor 4's "minor instances... that could be trimmed" better than anchor 5's "every token earns its place".

4 / 5

Actionability

Every event type has a fully executable, copy-paste-ready `opencastle log` command with realistic flag values, plus a concrete verification step (`tail -1 .opencastle/logs/events.ndjson`) and a field-level rule for `tier`/`model`. This matches anchor 5 ("Fully executable; copy-paste ready code or commands; specific examples cover the common cases"); anchor 4 would require minor gaps in the commands, and there are none.

5 / 5

Workflow Clarity

The sequence is explicit: an applicable event occurs → log it immediately ("One record per task; never batch-log retrospectively") → verify the append → confirm the pre-response checklist before responding, with "fix any missing log NOW" closing the feedback loop. The checklist and verification command are explicit validation checkpoints, matching anchor 5 ("explicit validation steps; feedback loops for error recovery; checklists") rather than anchor 4, whose checkpoints are only "mostly" present.

5 / 5

Progressive Disclosure

Good structure: a summary table up front, per-event-type command sections, a checklist, and a clearly signaled one-level reference ("See .opencastle/logs/README.md for full schema") that keeps bulk schema detail out of SKILL.md. It stops short of anchor 5 because the referenced README lives outside the skill bundle (no references/ directory exists in the skill, so the pointer cannot be resolved from the skill itself) and the body is slightly long for a pure overview. It clearly exceeds anchor 3, where references are buried or mis-scoped content sits inline.

4 / 5

Total

18

/

20

Passed

Description

88%Weight 40%Scale 1-5

Based on the skill's description, can an agent find and select it at the right time? Clear, specific descriptions lead to better discovery.

A strong description: third-person voice, concrete multi-action capability statement, and an explicit "Use when..." clause with enumerated trigger scenarios. The only weaknesses are a jargon-leaning trigger list with some generic head terms ("log", "track activity") that carry minor overlap risk and miss a few natural synonyms.

DimensionReasoningScore

Specificity

The description lists multiple concrete actions covering the domain comprehensively: "Logs sessions, tracks activity, records delegation decisions, and stores review/dispute outcomes as NDJSON audit trails" — these map directly to the five event types the skill handles. It clearly matches the anchor "Lists multiple specific concrete actions; comprehensive coverage" rather than anchor 4, which implies minor gaps; no capability area of the skill is left unnamed.

5 / 5

Completeness

It explicitly answers both questions: what ("Logs sessions, tracks activity, records delegation decisions, and stores review/dispute outcomes as NDJSON audit trails") and when ("Use when logging session activity, tracking work, recording decisions, building audit trails, capturing delegation history, or running pre-response verification checklists") with concrete trigger phrases. This matches the anchor-5 example pattern exactly; anchor 4 would require the 'when' to be less explicit or specific.

5 / 5

Trigger Term Quality

Trigger terms include natural phrases users would say ("log", "track activity", "audit trail", "session record") plus the explicit list. However, several terms lean toward internal jargon ("delegation log", "NDJSON", "pre-response verification checklists") and common variations or file extensions (e.g., "events.ndjson", "logging", "activity log") are missing. This fits anchor 4 ("Good keyword coverage; a few natural terms missing") better than anchor 5's comprehensive synonym/extension coverage, and better than anchor 3 since the listed terms are genuinely natural, not generic.

4 / 5

Distinctiveness Conflict Risk

The skill occupies a fairly distinct niche (audit-trail logging for agent delegation/review/panel workflows) with triggers unlikely to fire for unrelated skills. However, generic leading terms like "log" and "track activity" could overlap with other logging or task-tracking skills. This places it at anchor 4 ("Mostly distinct; minor overlap risk") rather than 5, which requires minimal conflict risk from every trigger.

4 / 5

Total

18

/

20

Passed

Validation

100%

Checks the skill against the spec for correct structure and formatting. All validation checks must pass before discovery and implementation can be scored.

Validation — 16 / 16 Passed

Validation for skill structure

No warnings or errors.

Repository
monkilabs/opencastle
Reviewed

Table of Contents

Is this your skill?

If you maintain this skill, you can claim it as your own. Once claimed, you can manage eval scenarios, bundle related skills, attach documentation or rules, and ensure cross-agent compatibility.