CtrlK
BlogDocsLog inGet started
Tessl Logo

langbot-testing

Test LangBot WebUI and core product flows with an automated browser and backend logs. Use when validating the configured LangBot frontend, pipeline Debug Chat, model provider setup and test buttons, bot and knowledge-base UI flows, or troubleshooting failed LangBot end-to-end tests.

71

Quality

86%

Does it follow best practices?

Run evals on this skill

Adds up to 20 points to the overall score

View guide

SecuritybySnyk

Low

Low-risk findings worth noting

SKILL.md
Quality
Evals
Security

Quality

Content

86%Weight 40%Scale 1-5

Reviews the quality of instructions and guidance provided to agents. Good implementation is clear, handles edge cases, and produces reliable results.

The body is a well-structured router skill: concise, actionable, with clear routing to verified reference files and explicit validation guardrails in the Rules section. It scores slightly below full marks on actionability and workflow_clarity because some runner-specific execution detail and explicit error-recovery feedback loops are deferred to the referenced files.

DimensionReasoningScore

Conciseness

Lean routing table and rules with no concept over-explanation; every line is actionable guidance or a guardrail, assuming Claude's competence and earning every token.

5 / 5

Actionability

Provides concrete executable commands ("bin/lbs fixture check", "bin/lbs suite start <suite-id>", env vars, specific reference files) with only minor gaps in runner-specific execution detail, matching the "mostly executable" anchor.

4 / 5

Workflow Clarity

The routing section sequences target selection and the rules include explicit validation checkpoints ("A WebUI test is not complete until... checked against backend logs", "Do not mark a case pass until test result --evidence covers...") for batch/evidence operations, with minor gaps in full error-recovery feedback loops.

4 / 5

Progressive Disclosure

Clear overview (Routing + Rules) pointing to 14 real one-level-deep reference files, each clearly signaled by scenario, with all referenced paths verified to exist; easy navigation and appropriate content split.

5 / 5

Total

18

/

20

Passed

Description

87%Weight 40%Scale 1-5

Based on the skill's description, can an agent find and select it at the right time? Clear, specific descriptions lead to better discovery.

The description is strong: it states concrete actions, provides an explicit "Use when" trigger covering multiple product flows, and carves out a distinct niche. The only minor gap is keyword synonym coverage, which keeps specificity and trigger_term_quality at 4 rather than 5.

DimensionReasoningScore

Specificity

Names the domain and several concrete actions ("Test LangBot WebUI...with an automated browser and backend logs") plus multiple product flow targets, with only minor coverage gaps, matching the "several specific actions" anchor.

4 / 5

Completeness

Explicitly answers both what (test WebUI and core flows via browser + backend logs) and when ("Use when validating the configured LangBot frontend, pipeline Debug Chat, model provider setup...") with concrete trigger phrases.

5 / 5

Trigger Term Quality

Good keyword coverage with natural phrases users would say ("WebUI", "Debug Chat", "model provider setup and test buttons", "knowledge-base UI flows", "troubleshooting failed end-to-end tests"), missing only a few synonyms.

4 / 5

Distinctiveness Conflict Risk

Clear LangBot product-testing niche with distinct triggers (WebUI, Debug Chat, Agent Runner, LangRAG, MCP ops) and minimal overlap risk with unrelated skills.

5 / 5

Total

18

/

20

Passed

Validation

100%

Checks the skill against the spec for correct structure and formatting. All validation checks must pass before discovery and implementation can be scored.

Validation16 / 16 Passed

Validation for skill structure

No warnings or errors.

Repository
langbot-app/LangBot
Reviewed

Table of Contents

Is this your skill?

If you maintain this skill, you can claim it as your own. Once claimed, you can manage eval scenarios, bundle related skills, attach documentation or rules, and ensure cross-agent compatibility.