CtrlK
BlogDocsLog inGet started
Tessl Logo

autogoal

Create, verify, repair, and close durable Codex goals with measurable outcomes, evidence gates, plan templates, blocker handling, completion audits, and goal-backed workflow repair.

68

Quality

83%

Does it follow best practices?

Run evals on this skill

Adds up to 20 points to the overall score

View guide

SecuritybySnyk

Passed

No findings from the security scan

SKILL.md
Quality
Evals
Security

Quality

Content

77%Weight 40%Scale 1-5

Reviews the quality of instructions and guidance provided to agents. Good implementation is clear, handles edge cases, and produces reliable results.

The body is an exceptionally actionable and well-sequenced policy document — copy-paste commands, explicit gates, conflict protocols, and validation loops throughout, all backed by real bundle files. Its weaknesses are length and shape: the same rules are restated multiple times, and nearly all policy lives inline in one ~1200-line file instead of being split across one-level-deep references.

Suggestions

Deduplicate the plan-editing rules (stated in Template Composition, Goal Plan creation, and Goal Plan editing) and the measurable-outcome rule (Core Take vs. Measurable Outcome Gate) into a single authoritative section, cutting repeated command blocks down to one canonical example.

Move self-contained policy blocks — Repair Mode with its scope matrix and workflow, Pass-Gated Goals, the template quality bar, and the required goal-plan section skeleton — into references/ files (e.g., references/repair-mode.md, references/template-quality.md), keeping SKILL.md as an overview with clearly signaled one-level-deep links.

Signal the bundle directly: link the bundled assets/templates/ templates and scripts/ helpers by their actual paths at the point of use, rather than only the installed .agents/skills/autogoal/... paths, so the relationship between SKILL.md and its bundle is navigable.

DimensionReasoningScore

Conciseness

The ~1200-line body is mostly novel repo-specific policy rather than explanations of known concepts, but the same rules are stated repeatedly: the fill-the-generated-plan/no-hand-narrowing rule appears in Template Composition, Goal Plan creation, and Goal Plan editing; "no measurable outcome, no goal" appears in both Core Take and the Measurable Outcome Gate; and the create-goal-scratchpad command block is shown three times. It is above anchor 2 because almost none of the content is padding about things Claude already knows, but below anchor 4 because a tightening pass could remove a substantial number of tokens.

3 / 5

Actionability

Guidance is fully executable: exact copy-paste commands with flags and paths (create-goal-scratchpad.mjs --template task --with docs, check-complete.mjs <path>, init-templates.mjs), a complete required goal-plan section skeleton, objective-handle templates, blocked/closeout report shapes, and explicit gate-table column specifications. All referenced scripts and template/pack files exist in the bundle, covering the common cases.

5 / 5

Workflow Clarity

The 15-step Start Workflow is clearly sequenced with explicit validation checkpoints (get_goal before create_goal, plan filled before substantive work, check-complete.mjs as the final mechanical gate before update_goal), plus an Active Goal Conflict Protocol, error-attempts tracking, a resume protocol, and validate→fix→retry loops (run scripts/validate-skills after editing skills, verify unfinished plans fail check-complete). This matches the anchor for clear sequences with explicit validation and feedback loops.

5 / 5

Progressive Disclosure

Section structure is strong and the bundle is real (scripts/ and assets/templates/ match every named template and pack), but the body is a ~1200-line monolith: Repair Mode, Pass-Gated Goals, the template quality bar, and the 70-line goal-plan skeleton clearly belong in reference files, and bundle files are referenced via installed .agents/skills/... paths rather than clearly signaled links to the bundled assets. Not a 2 because headers and section organization are good and external materials do exist; not a 4 because substantial content that should be split remains inline.

3 / 5

Total

16

/

20

Passed

Description

88%Weight 40%Scale 1-5

Based on the skill's description, can an agent find and select it at the right time? Clear, specific descriptions lead to better discovery.

A strong description: it explicitly covers the full goal lifecycle (create, verify, repair, close) with concrete supporting capabilities and an explicit, well-phrased use-when clause. The main gap is trigger phrasing — the most natural user phrases like "keep going until it works" or "set a goal" appear in the body but not the description, and one trigger leans on internal jargon ("governing repo skill").

DimensionReasoningScore

Specificity

The description lists multiple specific concrete actions spanning the full lifecycle — "Create, verify, repair, and close durable Codex goals" plus "measurable outcomes, evidence gates, plan templates, blocker handling, completion audits" — which matches the comprehensive-coverage anchor. It is not a 4 because the actions cover the whole goal lifecycle (create → verify → close, plus repair and blockers) rather than leaving minor gaps.

5 / 5

Completeness

It explicitly answers both questions: the first sentence states what the skill does (create/verify/repair/close goals with evidence gates and completion audits) and the second gives a concrete "Use when the user asks for a durable objective, long-running autonomous work, goal setup…" clause. This matches the anchor requiring both what and when with concrete trigger phrases.

5 / 5

Trigger Term Quality

Natural terms like "durable objective", "long-running autonomous work", and "goal setup" give good coverage, but common user phrasings the body itself relies on ("keep going", "set a goal", "continue") are absent, and "when a governing repo skill requires goal setup" is internal jargon. Not a 5 because several natural synonyms users would actually say are missing; not a 3 because the present keywords are genuinely relevant rather than generic.

4 / 5

Distinctiveness Conflict Risk

The Codex goal-lifecycle niche with evidence gates and completion audits is mostly distinct, but "long-running autonomous work" is broad enough to overlap with general autonomous-work or task-management skills. Not a 5 because of that breadth; not a 3 because the goal-tool framing (create_goal, blockers, completion audits) clearly separates it from neighboring skills.

4 / 5

Total

18

/

20

Passed

Validation

87%

Checks the skill against the spec for correct structure and formatting. All validation checks must pass before discovery and implementation can be scored.

Validation — 14 / 16 Passed

Validation for skill structure

CriteriaDescriptionResult

skill_md_line_count

SKILL.md is long (1207 lines); consider splitting into references/ and linking

Warning

referenced_paths_exist

Referenced path issues: 2 missing

Warning

Total

14

/

16

Passed

Repository
udecode/plate
Reviewed

Table of Contents

Is this your skill?

If you maintain this skill, you can claim it as your own. Once claimed, you can manage eval scenarios, bundle related skills, attach documentation or rules, and ensure cross-agent compatibility.