CtrlK
BlogDocsLog inGet started
Tessl Logo

grafana

Investigate production issues, query logs and metrics, and explore dashboards on the Mattermost Grafana instance at grafana.internal.mattermost.com.

64

Quality

78%

Does it follow best practices?

Run evals on this skill

Adds up to 20 points to the overall score

View guide

SecuritybySnyk

Passed

No findings from the security scan

Fix and improve this skill with Tessl

tessl review fix ./plugins/grafana/skills/grafana/SKILL.md
SKILL.md
Quality
Evals
Security

Quality

Content

90%Weight 40%Scale 1-5

Reviews the quality of instructions and guidance provided to agents. Good implementation is clear, handles edge cases, and produces reliable results.

A well-crafted, highly operational skill body: concrete and copy-paste-ready guidance, lean token usage, and a clear workflow with sensible query-strategy advice (targeted queries, namespace filtering, widening only if needed). The only real gaps are the absence of explicit validation checkpoints in the workflow and the lack of any progressive-disclosure structure given the skill slightly exceeds a single-screen length.

DimensionReasoningScore

Conciseness

The body is lean and efficient throughout: it assumes competence (never explains what Grafana, Loki, PromQL, or LogQL are), uses dense tables for UIDs, and every section adds non-obvious operational facts. The only trimmable redundancy is workflow step 2 restating the pre-cached UID instruction, which is a deliberate emphasis rather than padding.

5 / 5

Actionability

Guidance is fully executable: exact datasource UIDs, a concrete LogQL example ("{namespace=\"rxocmq9isjfm3dgyf4ujgnfz3c\"} |= \"error\""), parameter-by-parameter instructions for each tool, default time ranges, and specific defaults (limit 100). This matches the 'copy-paste ready, specific examples cover the common cases' anchor.

5 / 5

Workflow Clarity

The 6-step workflow (clarify scope → use cached UIDs → find dashboards → query → synthesize → deeplink) is clearly sequenced and includes one error-recovery hint ("Only call list_datasources if a query fails with an unknown UID error"), but there are no explicit validation checkpoints. This is a read-only investigation skill, so the destructive/batch cap does not apply; it sits above 'checkpoints missing or implicit' (3) but below the 'explicit validation steps; feedback loops' of a 5.

4 / 5

Progressive Disclosure

The skill is self-contained with no bundle files, and the body is well organized into clear sections (Environment, Namespace identifiers, per-tool querying, Workflow) with everything appropriately placed inline. At 68 lines it exceeds the 'under 50 lines with no external references → 5' exception, so it lands at 'good structure; minor organization gaps' rather than fully earning the top score.

4 / 5

Total

18

/

20

Passed

Description

66%Weight 40%Scale 1-5

Based on the skill's description, can an agent find and select it at the right time? Clear, specific descriptions lead to better discovery.

A solid, specific description with concrete actions and a clearly scoped niche, written in appropriate imperative/third-person style. Its main weakness is the absence of any 'Use when...' trigger guidance, which both caps completeness and slightly weakens trigger-term coverage and distinctiveness.

Suggestions

Append an explicit trigger clause, e.g. "Use when the user mentions Grafana, production issues, logs, metrics, dashboards, alerts, or oncall for Mattermost services."

Surface the additional capabilities already declared in allowed-tools (annotations, alert groups, oncall schedules) so the description's coverage matches the skill's actual scope.

Include common user phrasings such as "incident", "error rates", and "SLOs" to broaden natural trigger-term coverage.

DimensionReasoningScore

Specificity

The description names several concrete actions — "Investigate production issues, query logs and metrics, and explore dashboards" — scoped to a specific named instance, but omits capabilities reflected in allowed-tools (annotations, alerting, oncall), so coverage has minor gaps rather than being comprehensive.

4 / 5

Completeness

The 'what' is clear (investigate issues, query logs/metrics, explore dashboards on the Mattermost Grafana instance), but there is no 'Use when...' clause or equivalent explicit trigger guidance — only a weak implication via 'production issues'. Per judging guidelines a missing explicit 'when' caps this dimension at 3.

3 / 5

Trigger Term Quality

Natural keywords like "production issues", "logs", "metrics", "dashboards", and "Grafana" are present and would be said by users, but common variations such as "alerts", "incident", or "oncall" are missing. Fits the 'good keyword coverage; a few natural terms missing' anchor, not the comprehensive synonym/extension coverage of a 5.

4 / 5

Distinctiveness Conflict Risk

The description is anchored to a specific internal instance ("Mattermost Grafana instance at grafana.internal.mattermost.com"), giving it a clear niche distinct from generic monitoring skills. Minor overlap risk remains with general logging/observability skills since no explicit trigger phrases delineate when to prefer this one.

4 / 5

Total

15

/

20

Passed

Validation

93%

Checks the skill against the spec for correct structure and formatting. All validation checks must pass before discovery and implementation can be scored.

Validation — 15 / 16 Passed

Validation for skill structure

CriteriaDescriptionResult

allowed_tools_field

'allowed-tools' contains unusual tool name(s)

Warning

Total

15

/

16

Passed

Repository
mattermost/mattermost-ai-marketplace
Reviewed

Table of Contents

Is this your skill?

If you maintain this skill, you can claim it as your own. Once claimed, you can manage eval scenarios, bundle related skills, attach documentation or rules, and ensure cross-agent compatibility.