CtrlK
BlogDocsLog inGet started
Tessl Logo

evolve

하네스 진화 스킬. 사용 중인 하네스의 실행 결과에 대한 피드백을 수집·일반화하여 에이전트/스킬/오케스트레이터에 반영하고, 초기 구성 대비 델타를 포착해 변경 이력을 갱신한다. '하네스 회고', '하네스 진화', '하네스 피드백 반영', '하네스 개선', '결과가 아쉬웠어 하네스 고쳐줘', '이 피드백 하네스에 반영해줘', '하네스 레슨 정리' 등 기존 하네스의 실행 경험을 바탕으로 한 개선 요청 시 반드시 이 스킬을 사용. 하네스 신규 구축·구조 재설계·에이전트 추가는 harness 스킬이 담당.

76

Quality

95%

Does it follow best practices?

Run evals on this skill

Adds up to 20 points to the overall score

View guide
SecuritybySnyk

Passed

No findings from the security scan

SKILL.md
Quality
Evals
Security

Quality

Content

92%Weight 40%Scale 1-5

Reviews the quality of instructions and guidance provided to agents. Good implementation is clear, handles edge cases, and produces reliable results.

An exemplary instruction-only skill: a tight five-phase workflow with concrete commands, verbatim prompts, mapping tables, and a copy-paste output template, all validated with explicit checkpoints and regression guards. The body is lean and every section is actionable; only minor verification steps lack how-to detail.

Suggestions

Add a one-line example of what a 'should-trigger' vs 'near-miss' trigger test case looks like so the Phase 4 description verification is executable, not just named.

Give Phase 1's 산출물 품질 check a 2–3 item checklist (e.g., 형식 준수 / 경로 일치 / 기준 충족) to make dead-code detection mechanical.

DimensionReasoningScore

Conciseness

Lean and efficient throughout: no explanations of concepts Claude already knows, every section instructs rather than describes, and rationales ("왜 길어졌는지(스킬에 분량 배분 기준 부재)를 찾아") teach judgment instead of padding. The compact 원칙 section and single small diagram each earn their tokens.

5 / 5

Actionability

Highly concrete for an instruction skill: an exact command ("git log --oneline -- .claude/ CLAUDE.md"), verbatim user prompts, a feedback-type→target mapping table, and a copy-paste change-history table template with a filled example row. Not a 5 because a few steps stay directional — e.g., trigger verification is specified as "should-trigger + near-miss 각 3개 이상" without how to run it, and 산출물 품질 inspection has no checklist.

4 / 5

Workflow Clarity

Five clearly sequenced phases with explicit validation checkpoints in Phase 4 (structure check, trigger verification, CLAUDE.md consistency) plus error-recovery feedback loops: "변경은 한 번에 하나씩 적용하고, 각 변경 직후 Phase 4를 실행한다" and the 퇴행 방지 conflict check that pauses for user confirmation before reverting a past change.

5 / 5

Progressive Disclosure

No bundle files exist and none are needed: all content is single-file, well-organized with clear section headers, an overview diagram, and an explicit handoff to the separate harness skill for out-of-scope work. Nothing inlined here belongs in a separate reference file, and navigation via headers is trivial.

5 / 5

Total

19

/

20

Passed

Description

95%Weight 40%Scale 1-5

Based on the skill's description, can an agent find and select it at the right time? Clear, specific descriptions lead to better discovery.

A strong description: it states concrete third-person capabilities, provides an exhaustive set of natural Korean trigger phrases (formal and colloquial), explicitly mandates use conditions, and cleanly delimits scope from the sibling harness skill. The only minor gap is that a few workflow capabilities (reporting, trigger verification) are not surfaced.

Suggestions

Consider adding a brief mention of the 진화 보고 (evolution report) deliverable to round out the 'what' coverage without adding length.

DimensionReasoningScore

Specificity

Names the domain (harness evolution) and lists several concrete actions — "피드백을 수집·일반화하여 에이전트/스킬/오케스트레이터에 반영", "초기 구성 대비 델타를 포착해 변경 이력을 갱신" — in third person. Not a 5 because minor capabilities from the workflow (e.g., 진화 보고/validation of triggers) are absent, matching 'several specific actions; minor gaps'.

4 / 5

Completeness

Explicitly answers both: what ("피드백을 수집·일반화하여... 반영하고... 변경 이력을 갱신한다") and when ("...개선 요청 시 반드시 이 스킬을 사용") with concrete trigger phrases. Also adds an exclusion clause, so it exceeds the score-4 anchor ('when could be more explicit').

5 / 5

Trigger Term Quality

Comprehensive natural trigger coverage including synonyms and colloquial phrasings: "'하네스 회고'", "'하네스 진화'", "'하네스 피드백 반영'", "'하네스 개선'", "'결과가 아쉬웠어 하네스 고쳐줘'", "'이 피드백 하네스에 반영해줘'", "'하네스 레슨 정리'". These are exactly the phrases a user would naturally say; nothing common is missing.

5 / 5

Distinctiveness Conflict Risk

Clear niche with distinct triggers and an explicit boundary against the closest competing skill: "하네스 신규 구축·구조 재설계·에이전트 추가는 harness 스킬이 담당". Minimal conflict risk since all triggers are harness-evolution specific.

5 / 5

Total

19

/

20

Passed

Validation

100%

Checks the skill against the spec for correct structure and formatting. All validation checks must pass before discovery and implementation can be scored.

Validation — 16 / 16 Passed

Validation for skill structure

No warnings or errors.

Repository
revfactory/harness
Reviewed

Table of Contents

Is this your skill?

If you maintain this skill, you can claim it as your own. Once claimed, you can manage eval scenarios, bundle related skills, attach documentation or rules, and ensure cross-agent compatibility.