CtrlK
BlogDocsLog inGet started
Tessl Logo

ai-ide-runner

Run prompts in Claude Code, OpenCode, Cursor, or Codex CLIs from this session — one IDE, fan-out, or model comparison. Relay the other runtime's stdout verbatim. Use on run in <ide>, compare <ide> vs <ide>, try on <model>, run across models.

71

Quality

86%

Does it follow best practices?

Run evals on this skill

Adds up to 20 points to the overall score

View guide

SecuritybySnyk

Passed

No findings from the security scan

SKILL.md
Quality
Evals
Security

Quality

Content

85%Weight 40%Scale 1-5

Reviews the quality of instructions and guidance provided to agents. Good implementation is clear, handles edge cases, and produces reliable results.

The body is highly actionable with executable commands, a clear sequenced workflow, an explicit self-check validation checklist, and clean progressive disclosure to two real reference files. The main weakness is conciseness: the verbatim-relay/hook-blocked doctrine is repeated across several sections and could be consolidated.

Suggestions

Consolidate the repeated verbatim-relay/hook-blocked doctrine: the intro, 'Output contract', 'Hook-blocked tool calls', and 'Where the tool's stdout actually lives' all re-explain the same rule — merge them into one section with a single set of right/wrong examples.

Move the detailed blocked-run transcript examples (turn N / turn N+1 diagrams) into references/runtimes.md, keeping only the rule and one compact example inline in SKILL.md.

Tighten the 'Self-check before sending the final message' bullets into a compact checklist so the validation step reads as a checkpoint rather than three prose questions.

DimensionReasoningScore

Conciseness

The body is information-dense and avoids explaining basic concepts Claude already knows, but the 'courier, not a co-author' / hook-blocked / verbatim-relay doctrine is restated across the intro, 'Output contract', 'Hook-blocked tool calls', and 'Where the tool's stdout actually lives' sections with overlapping examples — noticeably more than minor trimming. It is not a 4 because the redundancy is repeated across multiple sections rather than isolated instances, and not a 2 because the content is genuinely useful and not padded with elementary explanations.

3 / 5

Actionability

Provides copy-paste-ready, fully executable commands for all four CLIs (e.g. `claude -p "<prompt>"`, `opencode run`, `cursor-agent -p --trust`, `codex exec`), a complete parallel `wait` block, and concrete fixes like `CLAUDECODE="" claude -p`. Placeholders for model IDs are justified by the explicit discovery guidance, so it is not a 4.

5 / 5

Workflow Clarity

The 4-step workflow (Parse intent → Pick model → Build & run → Present) is clearly sequenced, and the 'Self-check before sending the final message' section is an explicit validation checklist with error-recovery guidance (auth failures, native-first no-silent-fallback, missing binary). It is not a 4 because explicit validation steps and a checklist are present rather than merely implied.

5 / 5

Progressive Disclosure

The body is well-sectioned and offloads bulk detail to two real, one-level-deep, clearly-signaled reference files ([references/models.md], [references/runtimes.md]), keeping only a quick-start cheatsheet inline. It is not a 4 because navigation is clean and the reference split is appropriate rather than having only minor organization gaps.

5 / 5

Total

18

/

20

Passed

Description

87%Weight 40%Scale 1-5

Based on the skill's description, can an agent find and select it at the right time? Clear, specific descriptions lead to better discovery.

The description is strong: it states a concrete capability, gives explicit 'Use on...' trigger guidance covering single-IDE, fan-out, and model-comparison cases, and occupies a distinct niche. The only soft spot is trigger phrasing via placeholders rather than concrete synonyms.

DimensionReasoningScore

Specificity

Names the domain ('Run prompts in Claude Code, OpenCode, Cursor, or Codex CLIs') and several concrete actions (run one IDE, fan-out, model comparison, relay stdout verbatim), listing multiple specific actions with only minor coverage gaps. It is not a 5 because 'relay stdout verbatim' reads more as a constraint than a capability, and not a 3 since it clearly enumerates more than 1-2 actions.

4 / 5

Completeness

Explicitly answers both what ('Run prompts in ... CLIs ... Relay the other runtime's stdout verbatim') and when ('Use on run in <ide>, compare <ide> vs <ide>, try on <model>, run across models') with concrete trigger phrases. The 'Use on...' clause satisfies the explicit-trigger-guidance requirement, so it is not capped at 3 and clearly fits the 5 anchor.

5 / 5

Trigger Term Quality

Provides natural trigger phrasings via 'Use on run in <ide>, compare <ide> vs <ide>, try on <model>, run across models' alongside the named CLIs, giving good keyword coverage. It is not a 5 because the placeholders (<ide>, <model>) stand in for concrete synonyms and file extensions rather than enumerating them, and not a 3 because the trigger phrases are clearly natural user language.

4 / 5

Distinctiveness Conflict Risk

The niche — running prompts in four named IDE CLIs and relaying their output verbatim from this session — is highly distinct with minimal overlap risk against other skills. It is not a 4 because there is no meaningful overlap with closely related skills; the courier/verbatim framing carves out a clear, non-conflicting niche.

5 / 5

Total

18

/

20

Passed

Validation

100%

Checks the skill against the spec for correct structure and formatting. All validation checks must pass before discovery and implementation can be scored.

Validation16 / 16 Passed

Validation for skill structure

No warnings or errors.

Repository
korchasa/flowai-plugins
Reviewed

Table of Contents

Is this your skill?

If you maintain this skill, you can claim it as your own. Once claimed, you can manage eval scenarios, bundle related skills, attach documentation or rules, and ensure cross-agent compatibility.