CtrlK
BlogDocsLog inGet started
Tessl Logo

auto-review-loop-minimax

Autonomous multi-round research review loop using MiniMax API. Use when you want to use MiniMax instead of Codex MCP for external review. Trigger with "auto review loop minimax" or "minimax review".

61

Quality

73%

Does it follow best practices?

Run evals on this skill

Adds up to 20 points to the overall score

View guide

SecuritybySnyk

Low

Low-risk findings worth noting

Fix and improve this skill with Tessl

tessl review fix ./skills/skills-codex/auto-review-loop-minimax/SKILL.md
SKILL.md
Quality
Evals
Security

Quality

Content

63%Weight 40%Scale 1-5

Reviews the quality of instructions and guidance provided to agents. Good implementation is clear, handles edge cases, and produces reliable results.

The body delivers a genuinely actionable autonomous loop with concrete API calls, state persistence for compaction recovery, and an explicit stop condition. Its main costs are token efficiency (heavily duplicated prompt templates and meta-commentary) and progressive disclosure (prompt templates inlined rather than split into reference files, with no bundle present to back the shared-reference links).

Suggestions

Collapse the duplicated MCP and curl prompt blocks: define the system prompt and request shape once in 'API Configuration' and have 'Phase A' and 'Prompt Template for Round 2+' reference it instead of repeating them verbatim (~80 lines saved).

Move the full round-2+ prompt templates into a references/ file (e.g. references/prompts.md) and link to it, replacing the inline copies with a short summary of required sections (previous score/verdict/weaknesses, changes, updated results).

Add a concrete parsing/validation step in Phase B — e.g. extract the numeric score with a stated pattern, define fallback behavior when the reviewer response omits a score or verdict, and handle curl/API failures — to raise workflow clarity.

Delete the changelog meta-note about earlier 'or' wording and move the Codex Responses-API justification to a one-line link, keeping the operative stop condition only.

DimensionReasoningScore

Conciseness

The body is mostly operational, but the curl template and system prompt are duplicated nearly verbatim across 'API Configuration', 'Phase A', and 'Prompt Template for Round 2+' (~80 redundant lines), and it includes changelog meta-commentary ('Earlier wording used "or" + a stale verdict set; the AND form is authoritative.') and a justification paragraph with an external URL. This fits 'mostly efficient but could be tightened' better than level 2, since the core workflow content does earn its place.

3 / 5

Actionability

Concrete guidance throughout: copy-paste curl commands with headers and JSON body, exact MCP invocation, a full REVIEW_STATE.json schema, explicit file paths, prioritization rules, and a markdown template for documenting rounds. It stops short of level 5 because prompts rely on bracket placeholders and no concrete method is given for parsing score/verdict from the raw reviewer response or handling API errors.

4 / 5

Workflow Clarity

The sequence (Initialization, Phases A-E, Termination) is clearly laid out with an explicit STOP CONDITION, staleness/resume logic for state recovery, and a genuine review-fix-rereview feedback loop. It misses level 5 because there are no validation steps for failure modes: malformed or missing API responses, unparseable scores, or unavailable API keys are not addressed.

4 / 5

Progressive Disclosure

The body has reasonable section structure and a clearly signaled 'Output Protocols' block, but no bundle files exist (no references/, scripts/, or assets/), the linked shared-references resolve to paths outside the skill directory, and content that belongs in a separate reference file — the duplicated round-2+ prompt templates — is fully inlined. This matches 'some structure but could be better organized; content that should be separate is inline'.

3 / 5

Total

14

/

20

Passed

Description

82%Weight 40%Scale 1-5

Based on the skill's description, can an agent find and select it at the right time? Clear, specific descriptions lead to better discovery.

The description is well-constructed: it states what the skill does, when to use it, and how it differs from the Codex MCP alternative, with explicit trigger phrases. Its only notable weakness is specificity — it under-sells the loop's actual behavior (review, implement fixes, re-review) behind the abstract phrase 'review loop'.

DimensionReasoningScore

Specificity

The description names the domain ('Autonomous multi-round research review loop using MiniMax API') with one concrete action (multi-round review), but omits core capabilities performed in the body such as implementing fixes, re-reviewing, and persisting loop state. It matches the anchor 'names domain and 1-2 concrete actions, but not comprehensive' rather than level 4, which expects several specific listed actions.

3 / 5

Completeness

It explicitly answers both 'what' (autonomous multi-round research review loop via MiniMax API) and 'when' ('Use when you want to use MiniMax instead of Codex MCP for external review' plus concrete trigger phrases). This matches the top anchor, which requires both what and when with concrete trigger phrases; level 4 would apply only if the 'when' clause were weaker or implicit.

5 / 5

Trigger Term Quality

It provides explicit natural trigger phrases ('auto review loop minimax', 'minimax review') and the MiniMax/Codex MCP differentiator, giving good keyword coverage. It falls short of level 5 because common variations like 'external review', 'review my research', or 'research feedback' are missing.

4 / 5

Distinctiveness Conflict Risk

It carves out a clear niche (MiniMax API as the external reviewer instead of Codex MCP) with distinct, unambiguous trigger phrases, so the risk of firing for an unrelated review skill is minimal. This matches the 'clear niche with distinct triggers' anchor rather than level 4, which still assumes overlap with closely related skills.

5 / 5

Total

17

/

20

Passed

Validation

93%

Checks the skill against the spec for correct structure and formatting. All validation checks must pass before discovery and implementation can be scored.

Validation — 15 / 16 Passed

Validation for skill structure

CriteriaDescriptionResult

relative_links

Relative link issues: 3 suspicious

Warning

Total

15

/

16

Passed

Repository
wanshuiyin/Auto-claude-code-research-in-sleep
Reviewed

Table of Contents

Is this your skill?

If you maintain this skill, you can claim it as your own. Once claimed, you can manage eval scenarios, bundle related skills, attach documentation or rules, and ensure cross-agent compatibility.