CtrlK
BlogDocsLog inGet started
Tessl Logo

langbot-testing

Test LangBot WebUI and core product flows with an automated browser and backend logs. Use when validating the configured LangBot frontend, pipeline Debug Chat, model provider setup and test buttons, bot and knowledge-base UI flows, or troubleshooting failed LangBot end-to-end tests.

75

Quality

94%

Does it follow best practices?

Run evals on this skill

Adds up to 20 points to the overall score

View guide
SecuritybySnyk

Failed to scan

The risk profile of this skill

SKILL.md
Quality
Evals
Security

Quality

Content

93%Weight 40%Scale 1-5

Reviews the quality of instructions and guidance provided to agents. Good implementation is clear, handles edge cases, and produces reliable results.

An exemplary router-style SKILL.md: lean, fully actionable, with verified one-level-deep progressive disclosure and explicit validation gates for suites and evidence. The only soft spot is that the general workflow order is implied by rules rather than stated as a sequence.

DimensionReasoningScore

Conciseness

The body is a lean routing table of one-line pointers plus dense operational rules; it explains nothing Claude already knows and every line carries actionable information (env vars, commands, gates). Fits the 'lean and efficient; every token earns its place' anchor.

5 / 5

Actionability

Copy-paste-ready commands with concrete arguments ('bin/lbs fixture check', 'bin/lbs suite start <suite-id>', 'bin/lbs test result <case-id>', 'bin/lbs suite report <suite-id> --evidence-dir <dir>'), exact env var names, and exact reference paths. Per the code-vs-instruction note, the absence of code in this instruction-only router is not penalized.

5 / 5

Workflow Clarity

The suite lifecycle (suite start → per-case test result → suite report), preflight-before-gate ordering, and explicit validation gates ('Do not mark a case pass until test result --evidence covers every value in evidence_required', 'A WebUI test is not complete until the visible UI result is checked against backend logs') give clear checkpoints. Not 5 because the overall workflow ordering is implied by scattered rules rather than an explicit sequenced overview, and error-recovery loops live in troubleshooting.md.

4 / 5

Progressive Disclosure

A clean overview routing to 15 clearly signaled, one-level-deep reference files, all verified to exist, with detail appropriately split into the references (~1500 lines total) and the body kept minimal. Matches the 'clear overview with well-signaled one-level-deep references' anchor.

5 / 5

Total

19

/

20

Passed

Description

92%Weight 40%Scale 1-5

Based on the skill's description, can an agent find and select it at the right time? Clear, specific descriptions lead to better discovery.

A strong description with an explicit third-person 'what' and a concrete 'Use when' trigger clause covering the main LangBot test flows. Its only weakness is that roughly half of the skill's routed topic areas (plugins, MCP stdio, runners, performance/reliability, workspace release gates) are not represented as trigger terms.

Suggestions

Add one or two of the uncovered routing areas as natural trigger phrases (e.g., 'plugin install or runtime smoke tests', 'MCP stdio tool testing', 'runner release gates') so users asking about those flows match the skill via the description.

Consider a compact synonym phrase such as 'e2e / end-to-end tests' to catch both natural phrasings users say.

DimensionReasoningScore

Specificity

Enumerates multiple concrete actions ('Test LangBot WebUI and core product flows with an automated browser and backend logs', 'pipeline Debug Chat', 'model provider setup and test buttons', 'bot and knowledge-base UI flows', 'troubleshooting failed LangBot end-to-end tests') covering the skill's domain comprehensively; the 4 anchor would require minor coverage gaps, but the core flows are explicitly listed.

5 / 5

Completeness

Explicitly answers both 'what' (test LangBot WebUI and product flows with an automated browser and backend logs) and 'when' ('Use when validating the configured LangBot frontend... or troubleshooting failed LangBot end-to-end tests') with concrete trigger phrases, matching the 5 anchor.

5 / 5

Trigger Term Quality

Natural phrases users would say are present ('LangBot', 'WebUI', 'Debug Chat', 'test buttons', 'knowledge-base', 'end-to-end tests'), but several natural trigger areas from the skill's routing (plugin testing, MCP tools, runners, performance/reliability probes) are missing. Not 3 because the main use-case keywords are well covered; not 5 because coverage is not comprehensive.

4 / 5

Distinctiveness Conflict Risk

Clearly niche and LangBot-specific with named flows (pipeline Debug Chat, model provider test buttons, knowledge-base UI), so it is unlikely to trigger for the wrong skill; minimal overlap with generic testing skills.

5 / 5

Total

19

/

20

Passed

Validation

100%

Checks the skill against the spec for correct structure and formatting. All validation checks must pass before discovery and implementation can be scored.

Validation — 16 / 16 Passed

Validation for skill structure

No warnings or errors.

Repository
langbot-app/LangBot
Reviewed

Table of Contents

Is this your skill?

If you maintain this skill, you can claim it as your own. Once claimed, you can manage eval scenarios, bundle related skills, attach documentation or rules, and ensure cross-agent compatibility.