CtrlK
BlogDocsLog inGet started
Tessl Logo

remote-mac

Remote Macs: MacBooks, Mac Studios, hosted claw Macs, Tailscale, SSH, and OpenClaw.

53

Quality

59%

Does it follow best practices?

Run evals on this skill

Adds up to 20 points to the overall score

View guide

SecuritybySnyk

Passed

No findings from the security scan

Fix and improve this skill with Tessl

tessl review fix ./skills/remote-mac/SKILL.md
SKILL.md
Quality
Evals
Security

Quality

Content

71%Weight 40%Scale 1-5

Reviews the quality of instructions and guidance provided to agents. Good implementation is clear, handles edge cases, and produces reliable results.

The body is an unusually dense, high-signal operational runbook: concrete commands, explicit safety gates, and validation steps for risky operations, with almost no wasted tokens. Its main weaknesses are structural — everything lives in one monolithic file with no progressive-disclosure layer, plus some repetition and undated-section ephemeral state.

Suggestions

Split per-host topology and per-machine health checks into a references/ file (e.g., HOSTS.md, CHECKS.md) and keep SKILL.md as a concise overview with clearly signaled one-level-deep links.

Deduplicate repeated facts (Molty gateway retirement appears in both the topology list and the OpenClaw Checks section) and consolidate per-host state into one table.

Move time-sensitive incident state (2026-08-01 outages, ticket numbers, current regressed states) into a clearly labeled 'Current incidents' section so stale entries are easy to find and retire.

DimensionReasoningScore

Conciseness

The body is dense with fleet-specific facts and never explains concepts Claude already knows (e.g., 'GUI Tailscale builds cannot host Tailscale SSH', the exact ssh option flags). Minor trimmable redundancy exists — the Molty retirement is stated twice (topology and OpenClaw Checks sections) — and dated operational state ('Live 2026-08-01 state regressed') sits outside any deprecated/old-patterns section. Anchor 4, not 5.

4 / 5

Actionability

Concrete copy-paste commands cover the main cases: 'tailscale status --json', 'dns-sd -B _ssh._tcp local', 'lsof -nP -iTCP:18789 -sTCP:LISTEN', '~/.local/share/openclaw-clawstudio/run-current gateway status --deep --require-rpc --json', 'peekaboo list windows --app "Jump Desktop" --json'. Anchor 4 rather than 5 because some commands keep literal placeholders (HOST, COMMAND) and much per-host guidance is prose direction instead of executable form.

4 / 5

Workflow Clarity

Discovery is a numbered 1–8 sequence with verification steps ('Verify ComputerName, LocalHostName, hardware UUID'), and risky operations carry explicit checkpoints and gates ('verify with a raw window screenshot before clicking', 'One approval = one task', 'preserve a rollback path'), so the destructive-operation validation requirement is met. Anchor 4, not 5 because sections like Codex Automations and Live Testing Policy are rule lists rather than sequenced workflows with error-recovery loops.

4 / 5

Progressive Disclosure

The skill is a single ~130-line file with clear section headers but no bundle structure at all — no references/, scripts/, or assets/ exist, and per-host topology, GUI-access procedures, and per-machine health checks are all inlined where separate reference files would fit. The only pointers (computers.yaml, docs/tailnet-portal.md) target a private external manager repo, not navigable skill files. Anchor 3 (some structure, content that should be separate is inline), not 4 because there is no reference/navigation layer whatsoever.

3 / 5

Total

15

/

20

Passed

Description

47%Weight 40%Scale 1-5

Based on the skill's description, can an agent find and select it at the right time? Clear, specific descriptions lead to better discovery.

The description identifies a specific, distinguishable domain with good natural trigger keywords, but it says nothing about what the skill actually does and omits any 'Use when...' trigger guidance. It reads as a topic index rather than a capability statement, which caps both specificity and completeness low.

Suggestions

State concrete actions in the description (e.g., 'Discovers, SSHes into, runs health checks on, and manages OpenClaw gateways across Peter's remote Macs'), not just a list of host types and tools.

Add an explicit trigger clause: 'Use when the user says clawmac, megaclaw, miniclaw, clawstudio, Mac Studio, MacBook, Tailscale, or asks to run or check something on one of Peter's Macs.'

Include the specific hostnames users actually say (clawmac, megaclaw, miniclaw, foundationclaw) so the description's triggers match the body's trigger surface.

DimensionReasoningScore

Specificity

"Remote Macs: MacBooks, Mac Studios, hosted claw Macs, Tailscale, SSH, and OpenClaw" names the domain concretely (host types, tooling) but lists no actions or capabilities whatsoever. It matches anchor 2 (names the domain, actions minimal/generic) — not 3 because no concrete action is stated, not 1 because the domain is specifically enumerated.

2 / 5

Completeness

The description is a noun list with no verbs, so the 'what' is only a vague topic enumeration, and there is no 'Use when...' trigger clause at all. Anchor 2 (vague 'what', no 'when') — not 3 because even the 'what' lacks any concrete action; not 1 because the domain is at least named specifically.

2 / 5

Trigger Term Quality

Natural terms a user would say are present ("MacBooks", "Mac Studios", "Tailscale", "SSH", "OpenClaw", "Remote Macs"), but common variations the body itself relies on (clawmac, megaclaw, miniclaw, Mac Studio singular) are missing. Good coverage with a few natural terms absent — anchor 4, not 5.

4 / 5

Distinctiveness Conflict Risk

"hosted claw Macs" and "OpenClaw" carve a clear niche with distinct triggers, though "Remote Macs", "Tailscale", "SSH" are generic terms that overlap generic SSH/Tailscale or Mac-administration skills. Mostly distinct with minor overlap risk — anchor 4.

4 / 5

Total

12

/

20

Passed

Validation

100%

Checks the skill against the spec for correct structure and formatting. All validation checks must pass before discovery and implementation can be scored.

Validation — 16 / 16 Passed

Validation for skill structure

No warnings or errors.

Repository
steipete/agent-scripts
Reviewed

Table of Contents

Is this your skill?

If you maintain this skill, you can claim it as your own. Once claimed, you can manage eval scenarios, bundle related skills, attach documentation or rules, and ensure cross-agent compatibility.