CtrlK
BlogDocsLog inGet started
Tessl Logo

improve

Survey any codebase as a senior advisor and produce prioritized, self-contained implementation plans for OTHER models/agents to execute. Strictly read-only on source code — never implements, fixes, or refactors anything itself. Use when asked to audit a codebase, find improvement opportunities (bugs, security, performance, test coverage, tech debt, migrations, DX), suggest features or where to take the project next (roadmap, product direction), or generate handoff plans for another agent to implement.

76

Quality

94%

Does it follow best practices?

Run evals on this skill

Adds up to 20 points to the overall score

View guide

SecuritybySnyk

Low

Low-risk findings worth noting

The canonical home for this skill is improve in coder/agent-tty

SKILL.md
Quality
Evals
Security

Quality

Content

92%Weight 40%Scale 1-5

Reviews the quality of instructions and guidance provided to agents. Good implementation is clear, handles edge cases, and produces reliable results.

An exceptionally well-engineered skill body: a hard-rules contract, a phased workflow with vetting and verification gates baked in, executable commands and concrete output formats, and clean one-level-deep disclosure into three real, well-organized reference files. The only room for improvement is minor tightening of a few long, multi-clause passages.

DimensionReasoningScore

Conciseness

The body is dense and assumes Claude's competence (e.g. it never explains what a code audit or a git merge-base is), and every section carries operational weight — but a few passages could be trimmed: Hard Rule 1 packs three separate policies into one long sentence, and the "economics of this skill" paragraph is partly editorial. This fits the 4 anchor (efficient with minor over-explanation to trim) rather than 5, where every token would earn its place.

4 / 5

Actionability

Guidance is concrete and executable throughout: exact commands ("git rev-parse --short HEAD", "git diff --name-only $(git merge-base origin/<default> HEAD)..HEAD", "gh repo view --json visibility"), a concrete output directory layout, an enumerated subagent-prompt checklist, a specified findings-table schema, and named verification gates ("tsc --noEmit", lint check mode, audits). Per the rubric's code-vs-instruction note, an instruction-only skill with this level of specific, ready-to-run guidance matches the 5 anchor; it is not 4 because there are no significant gaps in executable detail.

5 / 5

Workflow Clarity

Four phases are explicitly sequenced (Recon → Audit → Vet/prioritize/confirm → Write plans) with validation checkpoints throughout: mandatory vetting of every subagent finding before it reaches the table, verification commands recorded in recon and stamped into every plan, drift detection against the recorded commit, escape hatches ("STOP and report back"), executor-diff review with hunk-level tracing, and reconciliation of prior runs. This matches the 5 anchor (clear sequence, explicit validation, feedback loops) and exceeds 4, where checkpoints would be only mostly present.

5 / 5

Progressive Disclosure

The body is a clear overview with three well-signaled, one-level-deep references that all exist and contain the cited sections: references/audit-playbook.md ("read it now", includes "## Finding format"), references/plan-template.md ("read it before writing the first plan"), and references/closing-the-loop.md (read "before the first dispatch"). Variant-specific detail is correctly deferred to those files rather than inlined. This matches the 5 anchor; it is not 4 because every reference is clearly signaled at the point of need and nothing that belongs in a bundle file is inlined.

5 / 5

Total

19

/

20

Passed

Description

96%Weight 40%Scale 1-5

Based on the skill's description, can an agent find and select it at the right time? Clear, specific descriptions lead to better discovery.

A strong description: concrete capabilities, third-person voice, an explicit 'Use when' clause with natural trigger phrasing, and a clear niche framed by the read-only/plans-for-other-agents boundary. The only weakness is moderate trigger-space overlap with general code-review and security-audit skills.

DimensionReasoningScore

Specificity

Multiple concrete actions are explicitly listed — "Survey any codebase", "produce prioritized, self-contained implementation plans", "audit a codebase", "find improvement opportunities (bugs, security, performance, test coverage, tech debt, migrations, DX)" — with comprehensive coverage of the audit domain and an explicit behavioral boundary ("Strictly read-only on source code — never implements, fixes, or refactors"). This matches the 5 anchor (multiple specific concrete actions, comprehensive); it is above 4 because there are no notable coverage gaps, and it is not below 5 since nothing listed is generic filler.

5 / 5

Completeness

Both halves are explicit: the "what" ("Survey any codebase as a senior advisor and produce prioritized, self-contained implementation plans for OTHER models/agents to execute") and the "when" ("Use when asked to audit a codebase, find improvement opportunities ..., suggest features ..., or generate handoff plans"). This matches the 5 anchor with concrete trigger phrases, exceeding 4 where the 'when' would be less specific.

5 / 5

Trigger Term Quality

Natural user phrasing is well covered with synonyms: "audit a codebase", "find improvement opportunities", "bugs, security, performance", "test coverage", "tech debt", "suggest features", "roadmap, product direction", "generate handoff plans". This fits the 5 anchor (comprehensive natural terms including synonyms) rather than 4, because the common variations users would actually say (audit / roadmap / features / handoff) are all present rather than just a few.

5 / 5

Distinctiveness Conflict Risk

The read-only advisor niche ("Strictly read-only on source code — never implements, fixes, or refactors anything itself"; plans written "for OTHER models/agents to execute") is clearly distinct from implement/refactor skills. It stays at 4 rather than 5 because the trigger space ("audit a codebase", "bugs, security, performance") overlaps with generic code-review / security-review skills, creating minor conflict risk with closely related skills — exactly the 4 anchor.

4 / 5

Total

19

/

20

Passed

Validation

100%

Checks the skill against the spec for correct structure and formatting. All validation checks must pass before discovery and implementation can be scored.

Validation — 16 / 16 Passed

Validation for skill structure

No warnings or errors.

Repository
openstatusHQ/data-table-filters
Reviewed

Table of Contents

Is this your skill?

If you maintain this skill, you can claim it as your own. Once claimed, you can manage eval scenarios, bundle related skills, attach documentation or rules, and ensure cross-agent compatibility.