CtrlK
BlogDocsLog inGet started
Tessl Logo

gstack-openclaw-retro

Weekly engineering retrospective. Analyzes commit history, work patterns, and code quality metrics with persistent history and trend tracking. Team-aware with per-person contributions, praise, and growth areas. Use when asked for weekly retro, what shipped this week, or engineering retrospective.

70

Quality

85%

Does it follow best practices?

Run evals on this skill

Adds up to 20 points to the overall score

View guide

SecuritybySnyk

Low

Low-risk findings worth noting

SKILL.md
Quality
Evals
Security

Quality

Content

82%Weight 40%Scale 1-5

Reviews the quality of instructions and guidance provided to agents. Good implementation is clear, handles edge cases, and produces reliable results.

A highly actionable, well-sequenced skill that trusts the reader with copy-paste git commands and concrete numeric thresholds throughout. Its main gaps are minor: a few editorial asides that pad token cost, no mid-flow validation checkpoints for partial data failures, and a single-file structure that inlines some material (output templates, teammate-analysis rules) that could be split into a reference file.

Suggestions

Add a brief validation checkpoint after Step 1 (e.g., confirm the fetch succeeded and the window returned commits before computing metrics) to strengthen the feedback loop.

Trim editorial commentary such as the 'ship fast, fix fast' explanation and keep rule statements bare.

Consider moving the Telegram output templates and per-teammate analysis rules into a single reference file to reduce always-loaded tokens.

DimensionReasoningScore

Conciseness

The body is command-first and dense: exact git format strings, numeric thresholds, and compact output templates, with no explanation of git concepts Claude already knows. A few editorial asides ('signals a ship fast, fix fast pattern that may indicate review gaps') could be trimmed, placing it at the 4 anchor rather than the lean 5.

4 / 5

Actionability

Every step is executable: copy-paste git commands with explicit format strings ('git log origin/main --since="<window>" --format="%H|%aN|%ae|%ai|%s" --shortstat'), concrete thresholds (45-minute session gap, PR buckets at 100/500/1500 LOC), and exact output formats. It fully matches the 5 anchor with specific examples covering the common cases.

5 / 5

Workflow Clarity

Fourteen clearly sequenced steps with explicit argument parsing, conditional branches (solo repo, window >= 14d), and terminal statuses (DONE / DONE_WITH_CONCERNS / BLOCKED) covering the not-in-a-repo and no-commits cases. It is not 5 because there are no explicit mid-flow checkpoints for partial data failures, e.g., verifying the initial fetch succeeded or handling a window with only some commands returning data.

4 / 5

Progressive Disclosure

The skill is a single well-sectioned file: no bundle directories exist and no external files are referenced, so there are no broken or nested references. It is not 5 because at ~300 lines some self-contained material (the Telegram output templates, per-teammate analysis rules) could plausibly live in a one-level-deep reference file to slim the always-loaded body.

4 / 5

Total

17

/

20

Passed

Description

87%Weight 40%Scale 1-5

Based on the skill's description, can an agent find and select it at the right time? Clear, specific descriptions lead to better discovery.

A strong description that clearly and explicitly states both what the skill does and when to use it, with natural trigger phrases and a distinctive niche. The only weakness is that capability coverage is stated at category level (work patterns, code quality metrics) rather than naming the concrete outputs the skill actually produces.

Suggestions

Name one or two signature outputs (e.g., per-contributor leaderboard, session detection, streak tracking) to sharpen the 'what' beyond category-level capability statements.

Add common trigger synonyms such as 'sprint retro', 'weekly recap', or 'what did we ship' to broaden natural keyword coverage.

DimensionReasoningScore

Specificity

The description lists several specific capabilities ('Analyzes commit history, work patterns, and code quality metrics', 'persistent history and trend tracking', 'per-person contributions, praise, and growth areas') but stays at the category level rather than naming concrete outputs such as leaderboards, session detection, or streak tracking. It lists several specific actions with minor gaps in coverage, matching the 4 anchor rather than the comprehensive 5.

4 / 5

Completeness

It explicitly answers both questions: what ('Analyzes commit history, work patterns, and code quality metrics with persistent history and trend tracking. Team-aware...') and when ('Use when asked for weekly retro, what shipped this week, or engineering retrospective') with concrete trigger phrases. This clearly matches the 5 anchor; it is not 4 because the 'when' clause is already explicit and specific.

5 / 5

Trigger Term Quality

'Use when asked for weekly retro, what shipped this week, or engineering retrospective' provides natural phrases a user would actually say. It sits at 4 rather than 5 because common variations like 'sprint retro', 'weekly recap', or 'what did we ship' are missing.

4 / 5

Distinctiveness Conflict Risk

'Weekly engineering retrospective' carves a clear niche with distinct triggers that would not naturally fire for adjacent git skills (commit messages, code review). Conflict risk with generic git-history skills is minimal, matching the 5 anchor.

5 / 5

Total

18

/

20

Passed

Validation

100%

Checks the skill against the spec for correct structure and formatting. All validation checks must pass before discovery and implementation can be scored.

Validation — 16 / 16 Passed

Validation for skill structure

No warnings or errors.

Repository
garrytan/gstack
Reviewed

Table of Contents

Is this your skill?

If you maintain this skill, you can claim it as your own. Once claimed, you can manage eval scenarios, bundle related skills, attach documentation or rules, and ensure cross-agent compatibility.