CtrlK
BlogDocsLog inGet started
Tessl Logo

add-model

Add a new LLM model to apps/sim/providers/models.ts with specs verified against the provider's live API docs (no hallucination)

64

Quality

76%

Does it follow best practices?

Run evals on this skill

Adds up to 20 points to the overall score

View guide

SecuritybySnyk

Low

Low-risk findings worth noting

Fix and improve this skill with Tessl

tessl review fix ./.agents/skills/add-model/SKILL.md
SKILL.md
Quality
Evals
Security

Quality

Content

88%Weight 40%Scale 1-5

Reviews the quality of instructions and guidance provided to agents. Good implementation is clear, handles edge cases, and produces reliable results.

An exemplary operational skill: hard anti-hallucination rules, a verified consumption matrix, exact commands and templates, and a mandatory verification-report format with escalation to the user. Its only real weakness is mild repetition of the hard rules across sections and an all-inline structure that forgoes any progressive disclosure.

DimensionReasoningScore

Conciseness

The body is dense with repo-specific facts Claude could not know (Consumption Matrix, hosted-billing rules, per-provider doc URLs) and wastes no tokens explaining known concepts, but there is deliberate redundancy — the Hard rules are restated across Steps 2, 4, and 6 ('Always re-grep before relying on this table', 'Cite every fact', 'never delete it to make lint pass') — a level of anchor-5 leanness it does not quite reach.

4 / 5

Actionability

Every instruction is copy-paste executable: exact WebFetch prompts, exact 'rg "reasoningEffort|reasoning_effort" apps/sim/providers/<provider>/' grep commands, a full TypeScript entry template with field-order and inline comments, exact 'bun run lint' / 'bun run agent-stream-docs:generate' commands, and a mandatory verification-report table format. Nothing is pseudocode or hand-waved.

5 / 5

Workflow Clarity

A clear 6-step sequence with explicit validation checkpoints and feedback loops: 'Lint must pass before you report done — fix the entry you wrote, never delete it to make lint pass', 'If you cannot reach an authoritative source for any field, mark the field as UNVERIFIED... ask the user before guessing', the hosted-billing assertion 'verify with getHostedModels().includes(...)', and the CI-diff check on regenerated docs. Error-recovery paths (source not found, lint failure, test breakage) are each explicitly handled.

5 / 5

Progressive Disclosure

Well-sectioned single-file skill with clear headings, tables, and step navigation, and no bundle files to mis-link. It sits at anchor 4 rather than 5 because ~190 lines are all inline with zero references — conditional sections like 'Reseller providers', 'Wrong family entirely?', and the full Consumption Matrix/URL table load on every invocation when they could live one level deep in reference files.

4 / 5

Total

18

/

20

Passed

Description

65%Weight 40%Scale 1-5

Based on the skill's description, can an agent find and select it at the right time? Clear, specific descriptions lead to better discovery.

A specific, distinct, third-person description anchored to a concrete file and a strong anti-hallucination guarantee. Its main gap is the missing 'when to use it' trigger clause, and secondarily it undersells the skill's actual scope (pricing verification, capability consumption checks, billing status).

Suggestions

Add an explicit trigger clause, e.g. 'Use when the user asks to add a new model, model pricing, or a model entry to the sim provider catalog' — this would lift completeness from 3 to 5.

Include natural synonyms users would say, such as 'model pricing', 'model list/catalog', or 'add <provider> model', to broaden trigger-term coverage.

Mention the pricing cross-check and capability/billing verification so the description covers what the skill actually does rather than only the insertion step.

DimensionReasoningScore

Specificity

'Add a new LLM model to apps/sim/providers/models.ts with specs verified against the provider's live API docs (no hallucination)' names the domain, one concrete action, and a concrete verification property with an exact file path — but it lists only a single action rather than the several specific actions (pricing cross-check, capability-flag consumption, hosted-billing verification) the skill actually performs, matching anchor 3 ('1-2 concrete actions, but not comprehensive') rather than anchor 4's 'several specific actions'.

3 / 5

Completeness

The 'what' is clear and concrete (add a model entry to a specific file, specs verified against live docs), but the 'when' is entirely absent — there is no 'Use when...' clause or equivalent trigger guidance, which per the judging guidelines caps completeness at 3.

3 / 5

Trigger Term Quality

'Add a new LLM model', 'provider', 'API docs', 'specs' are natural phrases a user would say when needing this skill, and 'no hallucination' captures the anti-drift intent. It stays at 4 rather than 5 because common variations a user might actually say — 'model pricing', 'new model entry', 'model catalog', 'model list', provider names — are absent.

4 / 5

Distinctiveness Conflict Risk

The repo-specific path 'apps/sim/providers/models.ts' plus the narrow action ('Add a new LLM model') gives it a clear niche with minimal overlap with generic documentation or coding skills — no other plausible skill would trigger on this description.

5 / 5

Total

15

/

20

Passed

Validation

93%

Checks the skill against the spec for correct structure and formatting. All validation checks must pass before discovery and implementation can be scored.

Validation — 15 / 16 Passed

Validation for skill structure

CriteriaDescriptionResult

frontmatter_unknown_keys

Unknown frontmatter key(s) found; consider removing or moving to metadata

Warning

Total

15

/

16

Passed

Repository
simstudioai/sim
Reviewed

Table of Contents

Is this your skill?

If you maintain this skill, you can claim it as your own. Once claimed, you can manage eval scenarios, bundle related skills, attach documentation or rules, and ensure cross-agent compatibility.