CtrlK
BlogDocsLog inGet started
Tessl Logo

observability

observability stuff

52

Quality

65%

Does it follow best practices?

Run evals on this skill

Adds up to 20 points to the overall score

View guide
SecuritybySnyk

Passed

No findings from the security scan

Fix and improve this skill with Tessl

tessl review fix ./observability/SKILL.md
SKILL.md
Quality
Evals
Security

Quality

Content

86%Weight 40%Scale 1-5

Reviews the quality of instructions and guidance provided to agents. Good implementation is clear, handles edge cases, and produces reliable results.

The body is a strong, tight instruction-only skill: it encodes specific team conventions Claude could not infer, is free of padding, and is well organized for its size. Its one real weakness is the Alerts section, which is placeholder-vague ("stuff is bad", "don't go overboard") compared to the concrete specificity of every other section.

DimensionReasoningScore

Conciseness

The body is lean throughout: it never explains what observability is or how the tools work, and every line encodes a team convention Claude could not guess (log levels, RED method, /metrics Prometheus format, OTel spans, 1%/100% sampling, required dashboard panels). This matches anchor 5 ("lean and efficient; assumes Claude's competence; every token earns its place"); the only soft wording ("stuff is bad") is two words and is an actionability issue, not padding.

5 / 5

Actionability

Most sections give concrete, executable guidance — "structured JSON logs at INFO/WARN/ERROR", "RED method... exposed at /metrics in Prometheus format", "sample traces... 1% baseline, 100% on error", specific dashboard panels — which meets anchor 4 ("mostly executable guidance... with minor gaps"). It falls short of anchor 5 because the Alerts section ("Just set up some alerts that page the team when stuff is bad. Don't go overboard.") gives no concrete thresholds, conditions, or examples — a genuine missing-key-detail gap, but localized to one of four sections, so anchor 3 ("incomplete") would be too harsh.

4 / 5

Workflow Clarity

The structure is a clear, well-ordered coverage sequence — three signals enumerated 1-2-3, gotchas, dashboards, alerts — and no destructive or batch operations exist, so no validation checkpoints are required. It matches anchor 4 ("clear sequence with most checkpoints present; minor validation gaps") rather than 5 because the single action is not fully unambiguous (the alerts step is undefined) and there is no verification step such as confirming /metrics responds or dashboards render.

4 / 5

Progressive Disclosure

The skill is well under 50 lines, has clear section headers, and contains no content that warrants splitting into reference files; no references/scripts/assets bundle exists, and the body references none. Per the rubric's simple-skill guideline, this earns anchor 5 ("well-organized sections" for a short skill with no need for external references).

5 / 5

Total

18

/

20

Passed

Description

28%Weight 40%Scale 1-5

Based on the skill's description, can an agent find and select it at the right time? Clear, specific descriptions lead to better discovery.

The description is a placeholder-level failure: two words with no capabilities, no triggers, and no use-when guidance. It names a real technical domain, which is the only thing keeping it from the bottom of the scale. The body of the skill contains everything needed to write a strong description, so this is an easy fix.

Suggestions

State concrete capabilities in the description, e.g. "Sets up structured JSON logging, RED metrics exposed at /metrics in Prometheus format, OpenTelemetry traces, and service dashboards with latency/error-rate panels."

Add an explicit trigger clause listing the natural terms users say, e.g. "Use when the user mentions observability, logging, metrics, traces, monitoring, dashboards, alerting, or instrumenting a service."

Keep third-person voice and drop the filler word "stuff" entirely — it adds no information and reads as an unfinished draft.

DimensionReasoningScore

Specificity

The description is only "observability stuff" — it names the domain (observability) but contains zero concrete actions; "stuff" is pure filler with no capability named. It sits between anchor 1 ("no concrete actions", e.g. "Helps with documents") and anchor 2 ("names the domain but actions are minimal or generic", e.g. "Processes PDF files"): naming a real technical domain keeps it above anchor 1, but unlike anchor 2 it offers not even a generic action verb.

2 / 5

Completeness

The "what" is present but vague ("observability stuff" says nothing about what is actually done) and there is no "when"/"Use when..." clause at all. Anchor 2 ("has a vague 'what' and no 'when'") is the best fit; it is above anchor 1 only because a domain-level "what" is at least stated.

2 / 5

Trigger Term Quality

The only keyword is "observability", which some users would say, but the natural phrases people actually use — logging, metrics, traces, monitoring, dashboards, alerting — are all absent. This matches anchor 2 ("one or two generic keywords; missing the natural phrases users say") rather than anchor 3, which requires a fuller set of relevant keywords.

2 / 5

Distinctiveness Conflict Risk

"observability" is a fairly specific niche term (not generic like "files" or "documents"), so it would not conflict with virtually any skill (anchor 1), but the undefined "stuff" scope gives it real overlap risk with any logging, monitoring, tracing, dashboarding, or alerting skill. Anchor 3 ("somewhat specific but could still overlap with similar skills") is the best fit; anchor 4 is not earned because nothing in the description delineates its scope from those neighbors.

3 / 5

Total

9

/

20

Passed

Validation

93%

Checks the skill against the spec for correct structure and formatting. All validation checks must pass before discovery and implementation can be scored.

Validation — 15 / 16 Passed

Validation for skill structure

CriteriaDescriptionResult

description_field

'description' is very short (19 chars), consider making it more detailed

Warning

Total

15

/

16

Passed

Repository
fernandezbaptiste/skill-review-sandbox
Reviewed

Table of Contents

Is this your skill?

If you maintain this skill, you can claim it as your own. Once claimed, you can manage eval scenarios, bundle related skills, attach documentation or rules, and ensure cross-agent compatibility.