CtrlK
BlogDocsLog inGet started
Tessl Logo

conductor-implement

Execute tasks from a track's implementation plan following TDD workflow

72

1.44x
Quality

68%

Does it follow best practices?

Impact

91%

1.44x

Average score across 3 eval scenarios

SecuritybySnyk

Passed

No findings from the security scan

Fix and improve this skill with Tessl

tessl review fix ./plugins/AI-Agents-Safe-Coding-Skills-claude/skills/conductor-implement/SKILL.md

The canonical home for this skill is conductor-implement in sickn33/agentic-awesome-skills

SKILL.md
Quality
Evals
Security

Quality

Content

70%Weight 40%Scale 1-5

Reviews the quality of instructions and guidance provided to agents. Good implementation is clear, handles edge cases, and produces reliable results.

The body is a well-sequenced, highly actionable implementation procedure with strong validation gates and error-recovery loops — its workflow clarity is exemplary. Its weaknesses are boilerplate filler sections, a loose test-command placeholder, and a single dangling reference to a nonexistent resources/implementation-playbook.md, leaving progressive disclosure underdeveloped.

Suggestions

Delete or replace the generic 'Use this skill when / Do not use this skill when / Instructions' filler sections with skill-specific content, tightening token usage.

Fix the progressive-disclosure gap: either create resources/implementation-playbook.md (e.g. moving the TDD phase details, error-handling menus, and completion templates there) or remove the dangling reference; clearly signal any reference files with a dedicated section.

Make the verification commands concrete — replace 'Run full test suite: npm test / pytest / etc.' with instructions to use the test command recorded in conductor/tech-stack.md or workflow.md.

DimensionReasoningScore

Conciseness

The core procedure (pre-flight, selection, task loop, completion) is dense and specific to the conductor system, but template-filler sections pad it out: 'Use this skill when: Working on implement track tasks or workflows', 'Do not use this skill when: The task is unrelated to implement track', and the generic Instructions bullets ('Clarify goals, constraints, and required inputs. Apply relevant best practices and validate outcomes.'). This fits anchor 3 ('Mostly efficient but includes some unnecessary explanation or could be tightened') — not a 2, since most of the body is genuinely non-redundant procedural content.

3 / 5

Actionability

The body gives exact file paths, copy-paste git commands ('git commit -m "{commit_prefix}: {task description} ({trackId})"'), concrete status-marker transitions ([ ] -> [~] -> [x]), exact menu templates, and a full metadata.json example. Minor gaps keep it below 5: the test command is loose ('Run full test suite: npm test / pytest / etc.') and test-writing guidance is abstract ('Write test(s) for the task functionality'). Anchor 4 ('Mostly executable guidance; concrete code or commands with minor gaps') is the best fit.

4 / 5

Workflow Clarity

The multi-step process is clearly sequenced with explicit validation checkpoints: TDD red/green/refactor with 'Run tests to confirm they fail' and HALT on unexpected passes, phase verification with 'CRITICAL: Wait for explicit user approval before proceeding to next phase', and final verification against spec.md acceptance criteria. Error handling provides explicit feedback loops (fix/rollback/pause options for tool, test, and git failures), matching anchor 5's 'explicit validation steps; feedback loops for error recovery'. No destructive/batch cap applies since validation is present throughout.

5 / 5

Progressive Disclosure

Section headers are clear, but the body is a ~390-line monolithic procedure with only one external reference — 'If detailed examples are required, open resources/implementation-playbook.md' — which is buried in the filler Instructions section and points to a file that does not exist (no references/, scripts/, or assets/ directories in the bundle). This fits anchor 3 ('Some structure but could be better organized; references present but not clearly signaled; content that should be separate is inline'): not a 2 because the inline content is one cohesive workflow with good headers, and not a 4 because the sole reference is unsignaled and dangling.

3 / 5

Total

15

/

20

Passed

Description

53%Weight 40%Scale 1-5

Based on the skill's description, can an agent find and select it at the right time? Clear, specific descriptions lead to better discovery.

The description states a clear, specific 'what' in third person but omits any 'when to use' trigger guidance, which limits discoverability. Keywords are domain-appropriate yet lack natural user phrasings and synonyms. Adding an explicit 'Use when...' clause with concrete triggers would resolve the main weakness.

Suggestions

Add a 'Use when...' clause with concrete triggers, e.g. 'Use when implementing tasks from a Conductor track plan, resuming paused tracks, or when the user asks to execute or continue a track.'

Enumerate a few more of the skill's concrete actions (per-task commits, phase verification gates, plan.md/metadata.json status updates) so the 'what' is comprehensive rather than minimal.

Include natural user phrasings and synonyms such as 'implement', 'build', 'continue working on', and 'track' next to the current 'Execute tasks' wording.

DimensionReasoningScore

Specificity

'Execute tasks from a track's implementation plan following TDD workflow' names the domain (track implementation plan, TDD) and two concrete actions (execute tasks, follow TDD workflow), but coverage is not comprehensive — the body's other capabilities (per-task commits, phase verification gates, status updates, resumption) are absent. This matches anchor 3 ('Names domain and 1-2 concrete actions, but not comprehensive') better than anchor 4, which expects several listed actions with only minor gaps.

3 / 5

Completeness

The 'what' is clear (execute implementation-plan tasks following TDD), but the 'when' is entirely missing — there is no 'Use when...' clause or equivalent trigger guidance, which caps completeness at 3 per the judging guidelines. It is not a 2 because the 'what' is specific rather than vague, and not a 4 because no explicit usage triggers exist at all.

3 / 5

Trigger Term Quality

Relevant keywords like 'implementation plan', 'TDD workflow', and 'tasks' appear, but common natural variations a user would actually say (e.g. 'implement', 'build the feature', 'continue working on the track') are missing. Anchor 3 ('Some relevant keywords but missing common variations or synonyms') fits; it is above anchor 2 (purely generic terms) and below anchor 4 (good coverage with few missing terms).

3 / 5

Distinctiveness Conflict Risk

'a track's implementation plan' plus 'TDD workflow' carves out a fairly distinct niche within the conductor workflow system, with only minor overlap risk against generic coding/TDD skills. This sits between anchor 3 ('could still overlap with similar skills') and anchor 5 (fully distinct triggers), fitting anchor 4 ('Mostly distinct; minor overlap risk with closely related skills').

4 / 5

Total

13

/

20

Passed

Validation

93%

Checks the skill against the spec for correct structure and formatting. All validation checks must pass before discovery and implementation can be scored.

Validation — 15 / 16 Passed

Validation for skill structure

CriteriaDescriptionResult

frontmatter_unknown_keys

Unknown frontmatter key(s) found; consider removing or moving to metadata

Warning

Total

15

/

16

Passed

Repository
administrakt0r/AI-Agents-Safe-Coding-Skills
Reviewed

Table of Contents

Is this your skill?

If you maintain this skill, you can claim it as your own. Once claimed, you can manage eval scenarios, bundle related skills, attach documentation or rules, and ensure cross-agent compatibility.