CtrlK
BlogDocsLog inGet started
Tessl Logo

connecting-to-sbx-sandboxes

Point an external editor or app — VS Code, Cursor, Claude Desktop, or ChatGPT — at a running Docker Sandbox over SSH via the `<name>.sbx` hostname, so the tool's UI stays on the host while files and processes run inside the sandbox. Use when setting up `sbx setup ssh`, wiring a remote-SSH connection in one of those apps, hitting a stalled or looping connection, or deciding whether Claude Desktop's credential-isolation warning applies to a given setup.

73

Quality

92%

Does it follow best practices?

Run evals on this skill

Adds up to 20 points to the overall score

View guide
SecuritybySnyk

Passed

No findings from the security scan

SKILL.md
Quality
Evals
Security

Quality

Content

88%Weight 40%Scale 1-5

Reviews the quality of instructions and guidance provided to agents. Good implementation is clear, handles edge cases, and produces reliable results.

A strong, dense body: every section carries sbx-specific knowledge Claude could not infer, commands are executable, and the workflow validates from a terminal before app wiring. The only weaknesses are inline version-gated claims (mitigated by a "Last verified" section) and per-app detail that could live in reference files.

Suggestions

Consolidate version-gated behavior changes (0.37.0/0.37.1/0.38.0) into a single version-history or "Last verified" section so inline claims don't read as time-sensitive assertions scattered through the body.

Move the per-app connection notes and the known-bugs table into a reference file (e.g. references/apps.md), keeping SKILL.md to the shared connect flow and the hostname mechanics.

DimensionReasoningScore

Conciseness

The body is dense and sbx-specific with essentially no explanation of concepts Claude already knows, but version-gated claims appear inline throughout ("Before 0.37.1", "0.37.1 (2026-07-29) turned this off by default", "From 0.38.0") and a few passages ("Same protocol, opposite purpose — don't let the shared vocabulary conflate them") could be tightened, fitting the efficient-with-minor-trimmings anchor at 4 rather than every-token-earns-its-place at 5.

4 / 5

Actionability

Guidance is copy-paste ready: `sbx setup ssh`, `sbx create --name demo shell .`, `ssh demo.sbx`, `sbx run --name <sandbox-name>`, `where.exe sh`, exact settings like `"remote.SSH.useLocalServer": false`, and a symptom-to-workaround table keyed on exact error strings — fully executable coverage of the common cases.

5 / 5

Workflow Clarity

The sequence is explicit with a validation checkpoint ("confirm `ssh <name>.sbx` works from a terminal first, then add the host inside the app's own remote/SSH UI") and a feedback loop via the known-bugs workaround table; no destructive or batch operations are involved, so no cap applies.

5 / 5

Progressive Disclosure

Sections are clearly organized and cross-skill routing (`running-sbx-sandboxes`, `diagnosing-sbx-sandboxes`) is one level deep, but at ~150 lines the per-app connection notes and bug table are inlined in SKILL.md with no bundle split, fitting the good-structure-with-minor-gaps anchor at 4 rather than the well-signaled-reference split at 5.

4 / 5

Total

18

/

20

Passed

Description

92%Weight 40%Scale 1-5

Based on the skill's description, can an agent find and select it at the right time? Clear, specific descriptions lead to better discovery.

An exemplary description: concrete actions, named target apps, an explicit "Use when..." clause with multiple concrete triggers, and a distinct niche that avoids conflict with sibling skills. The only minor gap is a few alternative natural phrasings a user might use.

DimensionReasoningScore

Specificity

The description lists multiple concrete actions — "Point an external editor or app... over SSH via the `<name>.sbx` hostname", "setting up `sbx setup ssh`", "wiring a remote-SSH connection", "hitting a stalled or looping connection", "deciding whether Claude Desktop's credential-isolation warning applies" — with named tools (VS Code, Cursor, Claude Desktop, ChatGPT) and no coverage gaps, matching the comprehensive-action anchor rather than the minor-gaps anchor at 4.

5 / 5

Completeness

It explicitly answers both what ("Point an external editor or app... at a running Docker Sandbox over SSH via the `<name>.sbx` hostname") and when ("Use when setting up `sbx setup ssh`, wiring a remote-SSH connection..., hitting a stalled or looping connection, or deciding whether Claude Desktop's credential-isolation warning applies"), matching the explicit what-and-when anchor at 5.

5 / 5

Trigger Term Quality

Natural trigger phrases are strong — app names, "SSH", "remote-SSH connection", "stalled or looping connection" — but a few phrasings a user might naturally say (e.g. "connect to sandbox", "remote development") are missing, fitting the good-coverage-with-a-few-gaps anchor at 4 rather than comprehensive synonym coverage at 5.

4 / 5

Distinctiveness Conflict Risk

It carves a clear niche — connecting external editors to Docker sbx sandboxes over the `<name>.sbx` SSH hostname — with triggers tied to `sbx setup ssh` and specific apps, so it is unlikely to fire for unrelated SSH or sandbox skills; fits the clear-niche anchor at 5.

5 / 5

Total

19

/

20

Passed

Validation

100%

Checks the skill against the spec for correct structure and formatting. All validation checks must pass before discovery and implementation can be scored.

Validation — 16 / 16 Passed

Validation for skill structure

No warnings or errors.

Repository
slurpyb/sbx-agent
Reviewed

Table of Contents

Is this your skill?

If you maintain this skill, you can claim it as your own. Once claimed, you can manage eval scenarios, bundle related skills, attach documentation or rules, and ensure cross-agent compatibility.