CtrlK
BlogDocsLog inGet started
Tessl Logo

livekit-simulations

Generate targeted test scenarios for a LiveKit voice or chat agent and run them as simulations — locally, from the agent's own code plus what the user wants stress-tested. Use whenever the user wants to "test my agent", "what should I test", "create/generate simulation scenarios", "make a sim test suite", "use lk agent simulate", "stress-test the X flow", "set up scenarios for my agent", or wants to probe edge cases / refusals / regressions before shipping. Generates scenarios on the user's machine (their code is never uploaded) and lets the user deeply steer what gets tested. Trigger even without the word "simulation" when the user clearly wants to decide what to test and verify how their agent behaves across realistic conversations. Not for building a new agent from scratch (use the livekit-agents skill), load-testing, or ordinary unit tests.

75

Quality

94%

Does it follow best practices?

Run evals on this skill

Adds up to 20 points to the overall score

View guide
SecuritybySnyk

Passed

No findings from the security scan

SKILL.md
Quality
Evals
Security

Quality

Content

85%Weight 40%Scale 1-5

Reviews the quality of instructions and guidance provided to agents. Good implementation is clear, handles edge cases, and produces reliable results.

A well-structured, actionable skill body that delegates detail to real reference files and scripts, with a clear validated workflow. The only slack is light editorial padding in the rationale sections.

Suggestions

Tighten the 'What makes this better than autopilot' and 'Principles' framing — the editorial justification ('A naive autopilot misses the point', 'The most valuable thing you can do') could be cut to pure directive guidance without losing meaning.

Consider moving the beta-notice block to a reference file or shortening it, since time-sensitive availability caveats add tokens that may churn before GA.

DimensionReasoningScore

Conciseness

Lean body that assumes Claude's competence and avoids explaining LiveKit or basic concepts, with only minor editorial framing ('The most valuable thing', 'A naive autopilot misses the point') that could be trimmed.

4 / 5

Actionability

Provides concrete executable commands with flags (build_scenarios.py assemble …, lk agent simulate --scenarios …) and named output files, with only the justified 'confirm exact flags with --help' hedge.

4 / 5

Workflow Clarity

A clear 5-step sequence with an explicit validation checkpoint in step 4 ('fails if any risk is uncovered … Fix gaps and re-run until it passes') and a feedback loop for error recovery, plus a reuse/regression note.

5 / 5

Progressive Disclosure

Body is a concise overview with well-signaled, one-level-deep references to real files (analyzing-the-agent.md, user-guidance.md, writing-scenarios.md) and a deterministic script, splitting detail appropriately.

5 / 5

Total

18

/

20

Passed

Description

100%Weight 40%Scale 1-5

Based on the skill's description, can an agent find and select it at the right time? Clear, specific descriptions lead to better discovery.

A strong, specific description with comprehensive trigger terms, explicit what/when guidance, and clear disambiguation against adjacent skills. No verbosity or vague fluff.

DimensionReasoningScore

Specificity

Lists multiple concrete actions — 'Generate targeted test scenarios', 'run them as simulations', 'stress-test the X flow', and 'probe edge cases / refusals / regressions' — with comprehensive coverage and no vague filler.

5 / 5

Completeness

Explicitly answers both 'what' (generate + run scenarios locally, never uploading code) and 'when' (concrete 'Use whenever' trigger phrases), and adds a 'Not for...' boundary clause.

5 / 5

Trigger Term Quality

Comprehensive natural phrases users would actually say — 'test my agent', 'what should I test', 'make a sim test suite', 'stress-test the X flow', 'set up scenarios for my agent' — plus synonyms and the CLI invocation 'lk agent simulate'.

5 / 5

Distinctiveness Conflict Risk

Clear niche (LiveKit agent simulations) with explicit disambiguation pointing to the livekit-agents skill and ruling out load-testing and unit tests, giving minimal conflict risk.

5 / 5

Total

20

/

20

Passed

Validation

100%

Checks the skill against the spec for correct structure and formatting. All validation checks must pass before discovery and implementation can be scored.

Validation — 16 / 16 Passed

Validation for skill structure

No warnings or errors.

Repository
livekit/agent-skills
Reviewed

Table of Contents

Is this your skill?

If you maintain this skill, you can claim it as your own. Once claimed, you can manage eval scenarios, bundle related skills, attach documentation or rules, and ensure cross-agent compatibility.