CtrlK
BlogDocsLog inGet started
Tessl Logo

testland/bug-bash-facilitator

Builds a structured bug-bash session - pre-bash kit (charter, test-data prep, environment setup, sign-up sheet), in-bash structure (role rotation across cohorts, shared backlog board, real-time triage), scoring rubric (severity weighting, novelty bonus), and a post-bash same-day wrap-up authored by the facilitator (not a standalone debrief: for post-session writeups without a live bash, use manual-test-debrief). Use when a team needs a coordinated multi-tester sweep before a release or after a major change - converts an ad-hoc "everyone test for an hour" into a recorded, comparable session with deliverables.

77

Quality

97%

Does it follow best practices?

Run evals on this skill

Adds up to 20 points to the overall score

View guide

SecuritybySnyk

Passed

No findings from the security scan

Overview
Quality
Evals
Security
Files

Quality

Content

92%

Reviews the quality of instructions and guidance provided to agents. Good implementation is clear, handles edge cases, and produces reliable results.

A dense, actionable facilitation playbook with copy-paste templates, a clearly sequenced timeline, and strong anti-pattern/limitations framing. Its only weakness is progressive disclosure: everything lives in one long SKILL.md with no external reference files, so the large pre-bash and debrief templates could be externalized for easier navigation.

Suggestions

Move the large Step 1 pre-bash kit template and Step 5 debrief template into separate reference files (e.g. references/pre-bash-kit-template.md and references/debrief-template.md), leaving concise overviews plus one-level-deep links in SKILL.md to improve progressive_disclosure.

Verify the three referenced skills (exploratory-tours-reference, manual-test-debrief, synthetic-data-tool-selector) exist in the bundle so the References links resolve to real files.

Consider trimming the fully-worked example rows in the Step 3 backlog and Step 5 debrief tables to blank/placeholder templates to further tighten conciseness.

DimensionReasoningScore

Conciseness

The body adds only domain-specific facilitation content (cohorts, tours, triage labels, scoring) and never explains concepts Claude already knows; it uses compact tables, checklists, and templates rather than prose padding, so every section earns its place. It is not the level below because there is no unnecessary explanation or generic library/background filler to tighten.

3 / 3

Actionability

Provides copy-paste-ready markdown templates for every phase — the pre-bash kit (Step 1), the time-boxed in-bash schedule (Step 2), the backlog-board column schema (Step 3), the points table (Step 4), the debrief template (Step 5), and the async mini-charter (Step 6) — with concrete repro examples and named triage labels. It is not the level below because the guidance is executable templates rather than pseudocode or vague direction.

3 / 3

Workflow Clarity

A clearly sequenced six-step process anchored to a timeline ('1 week before' → kickoff/mid-bash huddle/wrap → 'within 24 hours'), with checklists (test-data prep, action items) and explicit checkpoints like 'Verify staging is at the right SHA; tag artifact' and per-owner follow-up timeframes. It is not the level below because checkpoints and the sequence are explicit, not implicit; a bug-bash facilitation flow is also not a destructive/batch operation that would require a validate-fix-retry loop.

3 / 3

Progressive Disclosure

The skill is well-organized into clearly navigable sections (Overview, When to use, Steps 1-6, Anti-patterns, Limitations, References) and its References section clearly signals three related skills (exploratory-tours-reference, manual-test-debrief, synthetic-data-tool-selector), but no bundle files exist and the ~250-line body keeps all large templates inline with no file-level split. It is not the level above because nothing is appropriately split into one-level-deep reference files, and not the level below because it is well-sectioned with clearly signaled references rather than a monolithic wall or nested pointers.

2 / 3

Total

11

/

12

Passed

Description

100%

Based on the skill's description, can an agent find and select it at the right time? Clear, specific descriptions lead to better discovery.

A strong, third-person description that specifies concrete deliverables, includes an explicit 'Use when' trigger with natural terms, and proactively distinguishes itself from the related manual-test-debrief skill. All four dimensions hit the top anchor with no fluff or over-claims.

DimensionReasoningScore

Specificity

Lists multiple concrete actions across phases — 'pre-bash kit (charter, test-data prep, environment setup, sign-up sheet)', 'in-bash structure (role rotation across cohorts, shared backlog board, real-time triage)', 'scoring rubric (severity weighting, novelty bonus)', and 'post-bash same-day wrap-up' — matching the 'lists multiple specific concrete actions' anchor; not the level below because it is comprehensive rather than naming only a domain and some actions.

3 / 3

Completeness

Explicitly answers both what (the kit/structure/scoring/debrief components) and when via an explicit 'Use when a team needs a coordinated multi-tester sweep before a release or after a major change' clause; not the level below because the when-trigger is stated explicitly rather than merely implied.

3 / 3

Trigger Term Quality

Uses natural terms a team would actually say — 'bug-bash', 'coordinated multi-tester sweep', 'before a release or after a major change', and the ad-hoc phrasing 'everyone test for an hour' — giving good coverage of common variations; not the level below because it includes the everyday trigger phrasing rather than only narrow jargon.

3 / 3

Distinctiveness Conflict Risk

Occupies a clear niche (facilitated multi-tester bug bash) and actively disambiguates with 'not a standalone debrief: for post-session writeups without a live bash, use manual-test-debrief', so it is unlikely to trigger for the wrong skill; not the level below because the niche and boundary are explicit rather than merely 'somewhat specific'.

3 / 3

Total

12

/

12

Passed

Validation

100%

Checks the skill against the spec for correct structure and formatting. All validation checks must pass before discovery and implementation can be scored.

Validation16 / 16 Passed

Validation for skill structure

No warnings or errors.

Reviewed

Table of Contents