CtrlK
BlogDocsLog inGet started
Tessl Logo

langbot-plugin-dev

Develop, debug, and test LangBot plugins. Use when creating new LangBot plugins, fixing plugin bugs, setting up a LangBot test environment, or testing plugins via WebSocket. Covers plugin component architecture (EventListener, Command, Tool), the plugin SDK API (invoke_llm, get_llm_models, send_message, plugin storage), common pitfalls, and automated WebSocket-based testing. Triggers on "langbot plugin", "lbp", "GroupChatSummary", "plugin debug", "langbot test".

72

Quality

91%

Does it follow best practices?

Run evals on this skill

Adds up to 20 points to the overall score

View guide
SecuritybySnyk

Passed

No findings from the security scan

SKILL.md
Quality
Evals
Security

Quality

Content

78%Weight 40%Scale 1-5

Reviews the quality of instructions and guidance provided to agents. Good implementation is clear, handles edge cases, and produces reliable results.

A highly actionable, dense reference that earns its length with project-specific pitfalls, executable commands, and a strong debugging checklist. Main weaknesses are progressive disclosure (API reference, config types, and marketplace i18n conventions inlined instead of offloaded) and duplicated restart guidance with conflicting wait times.

Suggestions

Move the API Quick Reference (curl commands) and the README/i18n marketplace conventions into files under references/, keeping only the most-used commands inline to shrink the ~480-line body.

Merge "Plugin Hot-Reload" and "Container Restart Timing" into one section with a single consistent wait-time guidance (currently "~5 seconds" vs "~15 seconds").

Add an inline verification checkpoint to the test-environment workflow (e.g., "Verify plugin loaded: GET /api/v1/plugins" and confirm the model is configured before WebSocket testing) so failures are caught before test messages.

DimensionReasoningScore

Conciseness

Nearly all content is non-obvious project knowledge (SDK pitfalls with ❌/✅ pairs, event semantics, restart timing, trigger rules) with no padding about concepts Claude already knows. Not 5: "Plugin Hot-Reload" and "Container Restart Timing" duplicate the same restart guidance with inconsistent wait times ("~5 seconds" vs "~15 seconds"), and the marketplace README/i18n convention section runs long for inline placement.

4 / 5

Actionability

Copy-paste-ready curl commands with full JSON bodies, WebSocket URL templates with Origin-header code, executable ❌/✅ pitfall snippets, complete component YAML, and concrete docker restart sequences with timings. Not 4: the examples are fully executable and cover the common develop/test/debug/publish cases with no gaps.

5 / 5

Workflow Clarity

Setup, testing, debugging, and publishing are clearly sequenced, and the Debugging Checklist provides an explicit error-recovery loop (runtime logs → host logs → "Verify plugin loaded: GET /api/v1/plugins" → "Test person mode first" to isolate trigger rules). Not 5: the test-environment quick summary ends at "Copy plugin to data/plugins/" without inline verification checkpoints before WebSocket testing, relying on the debugging checklist for recovery.

4 / 5

Progressive Disclosure

The one bundle file (references/test-env-setup.md) is real and clearly signaled one level deep ("See references/test-env-setup.md for full deployment steps"), but substantial content that belongs in separate references is inlined: the full API curl Quick Reference, the Plugin Config Types table, and the README/i18n marketplace conventions push the body to ~480 lines. Not 4: more than minor organization gaps — several large sections are candidates for offloading while only one reference file exists.

3 / 5

Total

16

/

20

Passed

Description

100%Weight 40%Scale 1-5

Based on the skill's description, can an agent find and select it at the right time? Clear, specific descriptions lead to better discovery.

A model description: third-person, concrete actions with named SDK APIs, an explicit "Use when…" clause, and a dedicated trigger-term list covering the domain phrase, CLI name, and debug/test variants. Both what and when are answered without padding.

DimensionReasoningScore

Specificity

"Develop, debug, and test LangBot plugins" plus named coverage of "plugin component architecture (EventListener, Command, Tool), the plugin SDK API (invoke_llm, get_llm_models, send_message, plugin storage), common pitfalls, and automated WebSocket-based testing" lists multiple concrete actions with comprehensive domain coverage in third-person voice. Not 4: there are no coverage gaps — architecture, SDK API, pitfalls, and testing are each explicitly enumerated.

5 / 5

Completeness

Explicitly answers both "what" (develop/debug/test plus enumerated coverage areas) and "when" ("Use when creating new LangBot plugins, fixing plugin bugs, setting up a LangBot test environment, or testing plugins via WebSocket" plus concrete trigger phrases). Not 4: the "when" clause is already fully explicit and specific, not merely improvable.

5 / 5

Trigger Term Quality

Explicit triggers "langbot plugin", "lbp", "GroupChatSummary", "plugin debug", "langbot test" plus the "Use when creating… fixing plugin bugs… testing plugins via WebSocket" clause cover natural user phrasings, the CLI abbreviation, and variations. Not 4: synonym-level coverage (domain phrase, CLI name, debug/test variants, plugin-name trigger) is present rather than a few terms missing.

5 / 5

Distinctiveness Conflict Risk

"LangBot" names a distinct niche with unique triggers ("langbot plugin", "lbp") unlikely to fire for any other skill. Not 4: no closely-related-skill overlap is identifiable — the domain name and CLI abbreviation are proprietary to this ecosystem.

5 / 5

Total

20

/

20

Passed

Validation

100%

Checks the skill against the spec for correct structure and formatting. All validation checks must pass before discovery and implementation can be scored.

Validation — 16 / 16 Passed

Validation for skill structure

No warnings or errors.

Repository
langbot-app/LangBot
Reviewed

Table of Contents

Is this your skill?

If you maintain this skill, you can claim it as your own. Once claimed, you can manage eval scenarios, bundle related skills, attach documentation or rules, and ensure cross-agent compatibility.