CtrlK
BlogDocsLog inGet started
Tessl Logo

signals-scout-feature-flags

Signals scout for PostHog feature flags. Watches the flag roster and the `$feature_flag_called` stream for evaluation cliffs, ghost flags, response-distribution shifts, and flag debt, and files each validated contradiction as a report in the inbox.

67

Quality

81%

Does it follow best practices?

Run evals on this skill

Adds up to 20 points to the overall score

View guide

SecuritybySnyk

Low

Low-risk findings worth noting

SKILL.md
Quality
Evals
Security

Quality

Content

88%Weight 40%Scale 1-5

Reviews the quality of instructions and guidance provided to agents. Good implementation is clear, handles edge cases, and produces reliable results.

A highly actionable, well-sequenced skill body with strong validation checkpoints and clear structure; the only weakness is moderate length that could benefit from light tightening or external reference splitting.

DimensionReasoningScore

Conciseness

Dense, operational, domain-specific guidance with copy-paste SQL and no basic-concept padding, but at ~250 lines the introductory framing prose could be tightened in places.

4 / 5

Actionability

Fully executable guidance throughout — copy-paste-ready SQL queries, specific tool calls with parameters, and concrete scratchpad key patterns covering the common cases.

5 / 5

Workflow Clarity

A clearly sequenced run (orient → profile → explore → save memory → decide → close out) with explicit validation gates, volume gates, a disqualifier checklist, and an edit/author/remember/skip decision loop.

5 / 5

Progressive Disclosure

Well-organized ## and ### sections with the bulky report-channel contract correctly deferred to the harness prompt; no bundle files exist, but the body is long enough that some reference material could optionally be split out.

4 / 5

Total

18

/

20

Passed

Description

75%Weight 40%Scale 1-5

Based on the skill's description, can an agent find and select it at the right time? Clear, specific descriptions lead to better discovery.

A specific, well-scoped description with strong domain keywords and clear distinctiveness, but it lacks an explicit 'when to use' trigger clause, capping completeness.

Suggestions

Add a 'Use when...' clause naming the natural trigger phrases (e.g. 'Use when auditing PostHog feature flags for evaluation cliffs, ghost flags, or flag debt').

Include a few more colloquial synonyms alongside the technical terms to broaden trigger coverage.

DimensionReasoningScore

Specificity

Lists multiple concrete actions — 'Watches the flag roster and the `$feature_flag_called` stream for evaluation cliffs, ghost flags, response-distribution shifts, and flag debt, and files each validated contradiction as a report' — giving comprehensive coverage of what the scout does.

5 / 5

Completeness

The 'what' is clear and concrete, but there is no 'Use when...' clause or equivalent explicit trigger guidance, which per the judging guidelines caps completeness at 3.

3 / 5

Trigger Term Quality

Strong domain keywords a PostHog user would say ('feature flags', 'flag roster', 'ghost flags', 'flag debt', 'evaluation cliffs'), but it leans technical and omits common variations or an explicit natural trigger phrase.

4 / 5

Distinctiveness Conflict Risk

'Signals scout for PostHog feature flags' carves a clear niche with distinct triggers (feature flags, $feature_flag_called) and minimal overlap with sibling scouts.

5 / 5

Total

17

/

20

Passed

Validation

87%

Checks the skill against the spec for correct structure and formatting. All validation checks must pass before discovery and implementation can be scored.

Validation14 / 16 Passed

Validation for skill structure

CriteriaDescriptionResult

metadata_version

'metadata.version' is missing

Warning

frontmatter_unknown_keys

Unknown frontmatter key(s) found; consider removing or moving to metadata

Warning

Total

14

/

16

Passed

Repository
PostHog/posthog
Reviewed

Table of Contents

Is this your skill?

If you maintain this skill, you can claim it as your own. Once claimed, you can manage eval scenarios, bundle related skills, attach documentation or rules, and ensure cross-agent compatibility.