CtrlK
BlogDocsLog inGet started
Tessl Logo

debugging-livekit-agents

Drives a multi-turn conversation with a LiveKit agent running locally to see what it does. Use when the user says "test my agent", "try my agent", "does this work", "why did it call that tool", "it says the wrong thing when I ask X", "test this change", or whenever you have edited an agent and need to check how it behaves. Wraps `lk agent debugger`: start the agent in text mode, send turns, read the tool calls, handoffs, errors and logs behind each reply, and restart after an edit. It runs without audio or a LiveKit room at one LLM call per turn, so it is the preferred way for a coding agent to live-test during development, and the default when the user says "test" without naming unit tests or simulations.

76

Quality

96%

Does it follow best practices?

Run evals on this skill

Adds up to 20 points to the overall score

View guide
SecuritybySnyk

Passed

No findings from the security scan

SKILL.md
Quality
Evals
Security

Quality

Content

93%Weight 40%Scale 1-5

Reviews the quality of instructions and guidance provided to agents. Good implementation is clear, handles edge cases, and produces reliable results.

A lean, high-judgment body: executable commands for the main loop, debugging heuristics that are all non-obvious, a redirect table for when to use other tools, and deliberate non-duplication of `--help`. The only gap is a missing verification step that the debugger actually started before sending turns.

DimensionReasoningScore

Conciseness

Every sentence carries non-obvious domain knowledge (cost per turn, stale-code restart hazard, log lines interleaved under their turn) and nothing explains concepts Claude already knows. 'The help is thorough and is the source of truth for subcommands and flags, so this skill doesn't restate them' is exactly the right token economy.

5 / 5

Actionability

The core loop is copy-paste ready ('lk agent debugger start', 'lk agent debugger say "Hi, what can you do?"', 'lk agent debugger stop') and covers the common cases; withholding exact subcommand flags is explicitly justified as version-stability ('Check `--help` for the current format instead of assuming field names').

5 / 5

Workflow Clarity

'Start the agent, say user turns, inspect what happened, edit the code, restart' is a clear sequenced loop with a preflight check (confirm the command exists) and a reproduce-then-rerun feedback loop, but there is no explicit checkpoint that 'start' actually succeeded or what to do when a command itself errors — anchor 4's 'minor validation gaps' fits better than 5.

4 / 5

Progressive Disclosure

No bundle files exist and the ~75-line body is appropriately self-contained: well-organized sections (The loop, Debugging with it, When to use something else, Related skills), with detail delegated one level deep to `--help` and clearly named sibling skills in a decision table.

5 / 5

Total

19

/

20

Passed

Description

100%Weight 40%Scale 1-5

Based on the skill's description, can an agent find and select it at the right time? Clear, specific descriptions lead to better discovery.

A model description: concrete capability list, verbatim natural-language triggers, explicit what/when, and deliberate boundary marking against adjacent testing/simulation skills. It stays third person and every clause earns its place.

DimensionReasoningScore

Specificity

Quotes multiple concrete actions — 'start the agent in text mode, send turns, read the tool calls, handoffs, errors and logs behind each reply, and restart after an edit' — giving comprehensive coverage of the whole workflow rather than a generic domain label.

5 / 5

Completeness

Explicitly answers what ('Drives a multi-turn conversation... Wraps `lk agent debugger`', including the cost model) and when ('Use when the user says...') with concrete trigger phrases, matching the top anchor exactly.

5 / 5

Trigger Term Quality

Includes verbatim user utterances as triggers ('test my agent', 'try my agent', 'does this work', 'why did it call that tool', 'it says the wrong thing when I ask X', 'test this change') plus a catch-all behavioral trigger and a defaulting rule for the bare word 'test'.

5 / 5

Distinctiveness Conflict Risk

Clear LiveKit interactive-debugging niche with third-person voice; it explicitly carves itself away from unit tests and simulations ('the default when the user says "test" without naming unit tests or simulations'), leaving minimal conflict risk.

5 / 5

Total

20

/

20

Passed

Validation

93%

Checks the skill against the spec for correct structure and formatting. All validation checks must pass before discovery and implementation can be scored.

Validation — 15 / 16 Passed

Validation for skill structure

CriteriaDescriptionResult

metadata_version

'metadata.version' is missing

Warning

Total

15

/

16

Passed

Repository
livekit/agent-skills
Reviewed

Table of Contents

Is this your skill?

If you maintain this skill, you can claim it as your own. Once claimed, you can manage eval scenarios, bundle related skills, attach documentation or rules, and ensure cross-agent compatibility.