CtrlK
BlogDocsLog inGet started
Tessl Logo

drive-mimo

Use when developing or testing the MiMoCode repository and you need to programmatically drive another MiMoCode (mimo) process. Supports headless `mimo run` with JSON events and interactive TUI via tmux for behavior, integration, and visual regression testing. Covers an installed `mimo` binary or a dev build launched from source with `bun dev`; do not use it for ordinary MiMoCode task execution.

72

Quality

88%

Does it follow best practices?

Run evals on this skill

Adds up to 20 points to the overall score

View guide

SecuritybySnyk

Passed

No findings from the security scan

The canonical home for this skill is drive-mimo in XiaomiMiMo/MiMo-Code

SKILL.md
Quality
Evals
Security

Quality

Content

88%Weight 40%Scale 1-5

Reviews the quality of instructions and guidance provided to agents. Good implementation is clear, handles edge cases, and produces reliable results.

An excellent, highly actionable body: everything is executable, validation checkpoints are woven into every workflow, and the lone bundle script is clearly signaled. The deductions are minor — redundancy from the Quick Reference table, repeated Dev Mode facts, and inlining the scenario suite instead of splitting it into a separate file.

Suggestions

Trim the duplication: drop the Quick Reference table (it restates Part 1/Part 2 commands) or drop the inline equivalents, and consolidate the Dev Mode facts stated in both the Overview launcher table and the Dev Mode section.

Move the five scenario test functions (Part 3, ~90 lines) into a references/ or scripts/ file, keeping a short index in SKILL.md, so the main file reads as overview plus interface recipes.

Cut the tmux key explanations that restate standard tmux behavior (e.g., 'Up: History up', 'Tab: Tab completion') and keep only the mimo-specific usage.

DimensionReasoningScore

Conciseness

Mostly lean and dense with executable code and minimal prose — no padding explaining concepts Claude already knows — but there is measurable duplication: tmux key explanations ('Up: History up', 'Tab: Tab completion'), the Dev Mode launcher facts repeated across the Overview table, Dev Mode section, and substitution rule, and a Quick Reference table that restates body commands. Efficient with minor trimming opportunities, matching the 4 anchor rather than 5's 'every token earns its place'.

4 / 5

Actionability

Fully executable, copy-paste-ready commands throughout: launch snippets with exact flags ('MIMOCODE_HOME=$(mktemp -d) mimo run --format json --dangerously-skip-permissions'), jq/grep validation patterns, and five complete runnable scenario functions plus a batch runner. The referenced helper scripts/wait-for-text.sh exists with documented options. Matches the 5 anchor; concrete examples cover the common cases.

5 / 5

Workflow Clarity

Clear sequence (prerequisites → isolated launch → send input → wait-for-text → capture → cleanup) with explicit validation checkpoints everywhere: exit-code checks ('[ $EXIT -eq 0 ] || echo FAIL'), error-event greps, PASS/FAIL assertions per scenario, and a batch runner with a feedback summary. Even the batch operation has explicit verification, so the destructive/batch cap does not apply; this is the 5 anchor.

5 / 5

Progressive Disclosure

Good section structure (Overview, Dev Mode, Parts 1–3, Batch Runner, Quick Reference) and the single bundle file (scripts/wait-for-text.sh) is one level deep and well signaled with a callout explaining how to invoke it. Held at 4 rather than 5 because the ~450-line body inlines material that could be split out (the five scenario test functions, ~90 lines, are candidates for a references or scripts file) and the Quick Reference duplicates body content.

4 / 5

Total

18

/

20

Passed

Description

87%Weight 40%Scale 1-5

Based on the skill's description, can an agent find and select it at the right time? Clear, specific descriptions lead to better discovery.

A strong description with explicit what/when triggers, concrete capability enumeration, and an explicit exclusion boundary that minimizes conflict risk. Its only deductions are the second-person phrasing (specificity penalty) and a few missing natural trigger synonyms.

DimensionReasoningScore

Specificity

Lists multiple concrete actions ('headless `mimo run` with JSON events and interactive TUI via tmux for behavior, integration, and visual regression testing', dev build via `bun dev`) — comprehensive coverage that would merit a 5, but reduced by 1 because it uses second person ('you need to programmatically drive another MiMoCode (mimo) process') instead of third person.

4 / 5

Completeness

Explicitly answers both: what ('Supports headless `mimo run` with JSON events and interactive TUI via tmux') and when ('Use when developing or testing the MiMoCode repository'), plus a concrete do-not-use boundary ('do not use it for ordinary MiMoCode task execution'). Matches the 5 anchor exactly and is above the 4 anchor because the when-clause is explicit, not just present.

5 / 5

Trigger Term Quality

Good natural keyword coverage: 'developing or testing the MiMoCode repository', 'programmatically drive', 'mimo run', 'TUI via tmux', 'integration', 'visual regression testing', 'bun dev'. A few natural phrasings users might say are absent (e.g., 'automate mimo', 'spawn/test a mimo CLI session'), so it sits between the good-coverage (4) and comprehensive (5) anchors, closer to 4.

4 / 5

Distinctiveness Conflict Risk

Clear niche — programmatically driving a separate mimo process for testing — with an explicit negative trigger ('do not use it for ordinary MiMoCode task execution') that sharply reduces mis-triggering against general MiMoCode skills. Minimal conflict risk; clearly the 5 anchor over the 4 anchor's 'minor overlap risk'.

5 / 5

Total

18

/

20

Passed

Validation

100%

Checks the skill against the spec for correct structure and formatting. All validation checks must pass before discovery and implementation can be scored.

Validation — 16 / 16 Passed

Validation for skill structure

No warnings or errors.

Repository
XiaomiMiMo/MiMo-Code
Reviewed

Table of Contents

Is this your skill?

If you maintain this skill, you can claim it as your own. Once claimed, you can manage eval scenarios, bundle related skills, attach documentation or rules, and ensure cross-agent compatibility.