CtrlK
BlogDocsLog inGet started
Tessl Logo

agent-launcher-orchestrator

Use when a user wants to build, launch, grade, or schedule a Claude Managed Agent (CMA) in their own Anthropic account — "build me an agent", "launch this as a managed agent", "run this on a schedule", "grade my agent against a rubric", "set up a nightly worker". Reads the per-session goal (./my-agent/goal.json), routes deterministically to one of five phase sub-skills (interview → stage-launch → grade-iterate → run-without-you → wrap-up) via goal_router.py, and compiles the goal+phase into an execution shape (single-pass workflow / bounded grade→iterate loop / recurring cron deployment loop) via loop_compiler.py. Forks context so heavy intake (build sheets, payloads, eval cases) stays out of the parent thread. All launches are emitted as BYOK curl the user runs with their own key; no tool makes API calls. Inspired by anthropics/launch-your-agent (Apache-2.0). Distinct from engineering/agent-harness (generic domain loop) and engineering/write-a-skill (authors Claude Code skills, not CMAs).

74

Quality

93%

Does it follow best practices?

Run evals on this skill

Adds up to 20 points to the overall score

View guide
SecuritybySnyk

Passed

No findings from the security scan

The canonical home for this skill is agent-launcher-orchestrator in alirezarezvani/claude-skills

SKILL.md
Quality
Evals
Security

Quality

Content

88%Weight 40%Scale 1-5

Reviews the quality of instructions and guidance provided to agents. Good implementation is clear, handles edge cases, and produces reliable results.

A tight, well-sequenced orchestrator body: deterministic routing with exit-code semantics, explicit feedback loops, hard pre-flight refusals, and a compact tools index, with essentially no wasted tokens. The main weaknesses are the under-specified fork/hand-off step and references that point outside the skill directory and are not verifiable against the (absent) bundle.

Suggestions

Give the fork/hand-off as a concrete, copy-paste-ready step (e.g., the exact command or sub-skill invocation with the goal string, agent_name, out_dir, and plan.v1 arguments), so the whole route→compile→fork sequence is runnable end-to-end.

Fix reference paths so they resolve inside the skill bundle (e.g., references/cma-primitives.md rather than ../../references/...), and link 'interview-to-config.md' the same way as the other references.

Ensure the referenced files (cma-primitives.md, loops-and-workflows.md, interview-to-config.md) and scripts (goal_state.py, goal_router.py, loop_compiler.py) actually ship in the skill's references/ and scripts/ directories.

DimensionReasoningScore

Conciseness

The body is lean and assumes Claude's competence: short framed sections, a routing table, two compact command blocks, and a numbered forcing-question library — no explanation of concepts Claude already knows. The only trimmable bits are the provenance line ('Inspired by Anthropic's launch-your-agent reference skill (Apache-2.0)') and the opening paragraph echoing the description, which are minor; the content fits the 'every token earns its place' (score 5) anchor better than 'minor instances of over-explanation' (score 4).

5 / 5

Actionability

Mostly executable: copy-paste commands with documented exit-code semantics ('python3 scripts/goal_router.py --out-dir ./my-agent' with exit 0/3/4 meanings) and a fully flagged 'loop_compiler.py' invocation, plus concrete per-question recommendations and citations. It falls short of score 5 because the hand-off step ('fork to the sub-skill with: the goal string, agent_name, out_dir, and the compiled plan.v1') is described but not given as a runnable command, and no sub-skill invocation example is shown.

4 / 5

Workflow Clarity

The sequence is explicit and validated: read goal → run router → act on exit code → compile loop → fork to sub-skill → 'goal_state.py advance' → parent digest. Error-recovery feedback loops are explicit ('exit 3 ASK -> ask the one printed forcing question, then re-route'; 'exit 4 REFUSE -> ... get one sentence, then re-route'), and the 'Pre-flight gates (hard refusals)' section adds checkpoints ('Never route on under-3-word goals'). This matches the score-5 anchor (clear sequence, explicit validation, feedback loops); it is not score 4 because checkpoints are present throughout, not just most.

5 / 5

Progressive Disclosure

Good overview structure with clearly signaled references ('[`../../references/cma-primitives.md`](...)' and '[`../../references/loops-and-workflows.md`](...)') and a tools index — mostly matching the score-4 anchor. It is not score 5 because the referenced paths point outside the skill folder ('../../references/'), no bundle files exist alongside SKILL.md so the references cannot be verified to ship with the skill, and 'interview-to-config.md' is cited by bare name in the forcing-question library without a link or path.

4 / 5

Total

18

/

20

Passed

Description

100%Weight 40%Scale 1-5

Based on the skill's description, can an agent find and select it at the right time? Clear, specific descriptions lead to better discovery.

An exemplary description: explicit 'Use when' trigger clause with verbatim user phrasings, concrete and comprehensive action list in third person, and an explicit distinctness statement against neighboring skills. It is long, but every clause is load-bearing (triggers, mechanism, safety boundary, distinctness) rather than padded fluff.

DimensionReasoningScore

Specificity

Multiple concrete actions are explicitly listed: 'build, launch, grade, or schedule a Claude Managed Agent (CMA)', 'Reads the per-session goal (./my-agent/goal.json)', 'routes deterministically to one of five phase sub-skills', 'compiles the goal+phase into an execution shape', and 'All launches are emitted as BYOK curl'. This matches the score-5 anchor (multiple specific concrete actions, comprehensive coverage); it is not score 4 because coverage of the skill's behavior is complete rather than having minor gaps.

5 / 5

Completeness

Both 'what' ('Reads the per-session goal... routes deterministically to one of five phase sub-skills... compiles the goal+phase into an execution shape') and 'when' ('Use when a user wants to build, launch, grade, or schedule a Claude Managed Agent (CMA) in their own Anthropic account') are explicit and concrete with trigger phrases, exactly the score-5 anchor. It is not score 4 because the 'when' clause is fully explicit rather than improvable.

5 / 5

Trigger Term Quality

It embeds natural user phrases verbatim — '"build me an agent"', '"launch this as a managed agent"', '"run this on a schedule"', '"grade my agent against a rubric"', '"set up a nightly worker"' — plus 'Use when a user wants to...'. This comprehensively covers natural phrasings and their variations (the score-5 anchor), not merely a good subset (score 4).

5 / 5

Distinctiveness Conflict Risk

The niche is explicit — CMAs in the user's own Anthropic account — and it states 'Distinct from engineering/agent-harness (generic domain loop) and engineering/write-a-skill (authors Claude Code skills, not CMAs)', actively disambiguating against the closest neighbors. This is a clear niche with distinct triggers and minimal conflict risk (score 5), not merely 'mostly distinct' (score 4).

5 / 5

Total

20

/

20

Passed

Validation

81%

Checks the skill against the spec for correct structure and formatting. All validation checks must pass before discovery and implementation can be scored.

Validation — 13 / 16 Passed

Validation for skill structure

CriteriaDescriptionResult

frontmatter_unknown_keys

Unknown frontmatter key(s) found; consider removing or moving to metadata

Warning

relative_links

Relative link issues: 2 suspicious

Warning

referenced_paths_exist

Referenced path issues: 5 missing

Warning

Total

13

/

16

Passed

Repository
alirezarezvani/claude-skills
Reviewed

Table of Contents

Is this your skill?

If you maintain this skill, you can claim it as your own. Once claimed, you can manage eval scenarios, bundle related skills, attach documentation or rules, and ensure cross-agent compatibility.