CtrlK
BlogDocsLog inGet started
Tessl Logo

run-locally

Run and test the agent locally. Use when: (1) User says 'run locally', 'start server', 'test agent', or 'localhost', (2) Need curl commands to test API, (3) Troubleshooting local development issues, (4) Configuring server options like port or hot-reload.

71

Quality

86%

Does it follow best practices?

Run evals on this skill

Adds up to 20 points to the overall score

View guide

SecuritybySnyk

Passed

No findings from the security scan

The canonical home for this skill is run-locally in databricks/app-templates

SKILL.md
Quality
Evals
Security

Quality

Content

90%Weight 40%Scale 1-5

Reviews the quality of instructions and guidance provided to agents. Good implementation is clear, handles edge cases, and produces reliable results.

A tight, highly actionable reference skill with copy-paste commands and clear sections. Workflow clarity and progressive disclosure are strong but not maximal, as most sections are single-command references without explicit validation checkpoints or external file separation.

DimensionReasoningScore

Conciseness

The body is lean and code-forward — copy-paste commands with minimal supporting prose ('This starts the agent at http://localhost:8000', 'Uses MLflow scorers (RelevanceToQuery, Safety).') and no explanation of concepts Claude already knows, so every token earns its place per anchor 5.

5 / 5

Actionability

Provides fully executable, copy-paste-ready commands covering common cases (uv run start-app, full curl payloads with headers and JSON bodies, pytest, databricks experiments get-experiment), matching the anchor for fully executable guidance.

5 / 5

Workflow Clarity

Sections are clearly sequenced by task and the MLflow troubleshooting flow includes a verify-then-fix checkpoint, but most sections are single-command references without explicit validation feedback loops, fitting anchor 4 rather than the full-checklist 5.

4 / 5

Progressive Disclosure

Content is well-organized into clear sections (Start the Server, Server Options, Test the API, Troubleshooting, Next Steps) and appropriately self-contained with no nested references, but at ~85 lines with no external file split it sits at anchor 4 rather than the ideal one-level-deep reference structure of 5.

4 / 5

Total

18

/

20

Passed

Description

83%Weight 40%Scale 1-5

Based on the skill's description, can an agent find and select it at the right time? Clear, specific descriptions lead to better discovery.

A strong description that clearly states both capability and explicit 'Use when' triggers with natural user phrasing. Minor room for improvement in keyword synonym coverage and full action comprehensiveness.

DimensionReasoningScore

Specificity

Lists several concrete actions — 'Run and test the agent locally', curl commands to test the API, troubleshooting, and configuring server options like port or hot-reload — with only minor coverage gaps, fitting the 'several specific actions' anchor rather than the fully comprehensive 5.

4 / 5

Completeness

Explicitly states both what ('Run and test the agent locally') and when ('Use when: (1)... (2)... (3)... (4)...') with concrete trigger phrases, matching the anchor that requires clearly answering both what and when.

5 / 5

Trigger Term Quality

Includes natural phrases users would say ('run locally', 'start server', 'test agent', 'localhost') plus technical terms (curl, API, hot-reload), giving good keyword coverage though a few synonyms are missing, so it sits above anchor 3 but below the fully comprehensive 5.

4 / 5

Distinctiveness Conflict Risk

The local-dev/localhost niche is mostly distinct from sibling skills like deploy, with only minor overlap risk on closely related skills, fitting anchor 4 rather than the minimal-conflict 5.

4 / 5

Total

17

/

20

Passed

Validation

100%

Checks the skill against the spec for correct structure and formatting. All validation checks must pass before discovery and implementation can be scored.

Validation16 / 16 Passed

Validation for skill structure

No warnings or errors.

Repository
databricks/app-templates
Reviewed

Table of Contents

Is this your skill?

If you maintain this skill, you can claim it as your own. Once claimed, you can manage eval scenarios, bundle related skills, attach documentation or rules, and ensure cross-agent compatibility.