CtrlK
BlogDocsLog inGet started
Tessl Logo

llm-endpoint-demo

Demo and end-to-end fixture for MiMoCode's temporary local LLM server. Use when verifying that a skill can borrow a configured chat model through a local base_url and a throwaway token instead of a real provider API key, or when testing the expire-and-reissue loop. Not a general-purpose skill.

69

Quality

84%

Does it follow best practices?

Run evals on this skill

Adds up to 20 points to the overall score

View guide

SecuritybySnyk

Passed

No findings from the security scan

SKILL.md
Quality
Evals
Security

Quality

Content

81%Weight 40%Scale 1-5

Reviews the quality of instructions and guidance provided to agents. Good implementation is clear, handles edge cases, and produces reliable results.

A well-structured, actionable skill body with a clear setup workflow and an explicit error-recovery feedback loop for token expiry. It is efficient and concrete, held back only slightly by philosophical framing and copy-paste-blocking placeholders.

DimensionReasoningScore

Conciseness

Mostly lean and assumes Claude's competence (env-var table, direct commands); minor over-explanation in the 'black-box witness' framing and rationale prose keeps it just below a 5.

4 / 5

Actionability

Provides concrete, mostly executable commands (llm-server status/issue, export+run invocation) and an exit-code-to-action table; not a 5 because placeholders like <mimocode> and <provider/model> prevent full copy-paste readiness.

4 / 5

Workflow Clarity

Setup is clearly sequenced (steps 1-3) and the expiry section gives an explicit feedback loop (exit 2 -> run renew_argv -> replace key -> retry once) with a full exit-code recovery table.

5 / 5

Progressive Disclosure

Well-organized into clear sections (needs, setup, expiry, prohibitions) with all content appropriately inline and no nested references; not a 5 because the body exceeds the ~50-line simple-skill threshold and there is no reference structure to signal.

4 / 5

Total

17

/

20

Passed

Description

87%Weight 40%Scale 1-5

Based on the skill's description, can an agent find and select it at the right time? Clear, specific descriptions lead to better discovery.

A strong, niche-scoped description that clearly answers both what the skill does and when to use it, with concrete mechanisms and explicit non-general-purpose boundary. The only weak spot is keyword synonym/extension coverage, which is inherently limited for such a specialized demo.

DimensionReasoningScore

Specificity

Names the domain ('temporary local LLM server') and several concrete mechanisms ('borrow a configured chat model through a local base_url and a throwaway token', 'expire-and-reissue loop'), with only minor coverage gaps; not a 5 because it frames roughly two actions rather than a comprehensive list.

4 / 5

Completeness

It explicitly states both what it is ('Demo and end-to-end fixture for MiMoCode's temporary local LLM server') and when to use it ('Use when verifying... or when testing the expire-and-reissue loop') with concrete trigger phrases.

5 / 5

Trigger Term Quality

The 'Use when verifying that a skill can borrow a configured chat model... or when testing the expire-and-reissue loop' clause gives good keyword coverage for its narrow niche; not a 5 because it lacks synonyms or file-extension variants.

4 / 5

Distinctiveness Conflict Risk

Highly niche triggers ('MiMoCode's temporary local LLM server', 'expire-and-reissue loop') plus the explicit 'Not a general-purpose skill' boundary give it a clear niche with minimal conflict risk.

5 / 5

Total

18

/

20

Passed

Validation

100%

Checks the skill against the spec for correct structure and formatting. All validation checks must pass before discovery and implementation can be scored.

Validation — 16 / 16 Passed

Validation for skill structure

No warnings or errors.

Repository
XiaomiMiMo/MiMo-Code
Reviewed

Table of Contents

Is this your skill?

If you maintain this skill, you can claim it as your own. Once claimed, you can manage eval scenarios, bundle related skills, attach documentation or rules, and ensure cross-agent compatibility.