CtrlK
BlogDocsLog inGet started
Tessl Logo

rebuttal

Workflow 4: Submission rebuttal pipeline. Parses external reviews, enforces coverage and grounding, drafts a safe text-only rebuttal under venue limits, and manages follow-up rounds. Use when user says "rebuttal", "reply to reviewers", "ICML rebuttal", "OpenReview response", or wants to answer external reviews safely.

71

Quality

88%

Does it follow best practices?

Run evals on this skill

Adds up to 20 points to the overall score

View guide

SecuritybySnyk

Low

Low-risk findings worth noting

SKILL.md
Quality
Evals
Security

Quality

Content

77%Weight 40%Scale 1-5

Reviews the quality of instructions and guidance provided to agents. Good implementation is clear, handles edge cases, and produces reliable results.

An unusually rigorous, executable workflow with excellent sequencing, validation gates, and safety controls — the operational core is top-tier. Weaknesses are length/redundancy (several concepts defined twice) and progressive disclosure: ~370 lines inlined where templates and schemas belong in bundle files, plus citations to shared-references files that are absent from the bundle.

Suggestions

Ship the referenced bundle files or inline their content: 'shared-references/reviewer-routing.md', 'shared-references/review-tracing.md', 'shared-references/integration-contract.md', and 'save_trace.sh' are cited but absent from the bundle, so review tracing (Policy C) is currently unexecutable.

Move stable schemas and templates out of SKILL.md into references/ (issue-card field definitions, the REVISION_PLAN checklist template, and the reviewer calling convention configs) to cut the body from ~370 lines to an overview.

Deduplicate repeated definitions: 'structural_distinction' (Phase 2 vs. Reviewer-defensive moves), pivotal reviewer allocation (Constants vs. Phase 3), and the VENUE_MODE output structure (Phase 4 vs. Phase 7) are each explained twice — keep one canonical definition.

DimensionReasoningScore

Conciseness

Mostly efficient operational detail ("Sentence 1: direct answer / Sentence 2-4: grounded evidence"), but with noticeable duplication: pivotal reviewers are defined in both the Constants ("STRESS_TEST_ROUNDS_BASE") and Phase 3 step 4; 'structural_distinction' is explained in full twice (Phase 2 response_mode list and 'Reviewer-defensive moves'); the VENUE_MODE draft structure is stated in Phase 4 and restated in Phase 7; REVISION_PLAN rules appear in both Phase 4 and the checklist rationale. This fits 'mostly efficient but includes some unnecessary explanation or could be tightened' rather than the minor-trim profile of a 4.

3 / 5

Actionability

Fully executable guidance: exact MCP call templates with configs ("config: {\"model_reasoning_effort\": \"xhigh\"}"), a copy-paste stress-test prompt, exact artifact paths (`rebuttal/ISSUE_BOARD.md`, `rebuttal/PASTE_READY.txt`), a concrete markdown checklist example with issue_id/commitment/status fields, and explicit slash-command invocations ("/experiment-bridge \"rebuttal/REBUTTAL_EXPERIMENT_PLAN.md\""). Matches 'copy-paste ready; specific examples cover the common cases'. Not a 4 because the few abstract steps ('pause and ask') are intentional control flow, not missing detail.

5 / 5

Workflow Clarity

Phases 0-9 are clearly sequenced with resume handling ("If rebuttal/REBUTTAL_STATE.md exists → resume from recorded phase"), an explicit validation phase (Phase 5's eight lints: coverage, provenance, commitment, tone, consistency, limit, thread-local context, adversarial scan), and real feedback loops ("If any hard safety blocker remains → revise before finalizing", re-run lints in follow-up rounds). This is the anchor-5 profile of sequence + explicit validation + error-recovery loops.

5 / 5

Progressive Disclosure

The body has good section structure, but the bundle contains no references/, scripts/, or assets/ directories while the text cites "shared-references/reviewer-routing.md", "shared-references/review-tracing.md", "shared-references/integration-contract.md §2", and "save_trace.sh" — dangling references to files not shipped. Additionally, substantial content that belongs in reference files is inlined (issue-card field schema, REVISION_PLAN template, reviewer calling conventions). Fits 'references present but not clearly signaled; content that should be separate is inline'. Not a 2 because the body itself is well-organized with navigable headers rather than a reference-buried or structureless wall.

3 / 5

Total

16

/

20

Passed

Description

100%Weight 40%Scale 1-5

Based on the skill's description, can an agent find and select it at the right time? Clear, specific descriptions lead to better discovery.

A model description: third-person voice, concrete multi-action capability summary, and an explicit 'Use when...' clause enumerating natural trigger phrases including venue-specific synonyms. No over-claims or padding.

DimensionReasoningScore

Specificity

The description lists four concrete, distinct actions — "Parses external reviews", "enforces coverage and grounding", "drafts a safe text-only rebuttal under venue limits", "manages follow-up rounds" — covering the full pipeline comprehensively, matching the anchor 'multiple specific concrete actions; comprehensive coverage'. Not a 4 because no meaningful capability of the pipeline is omitted from the summary.

5 / 5

Completeness

Explicitly answers both questions: 'what' via the four action clauses and 'when' via "Use when user says 'rebuttal', 'reply to reviewers', ..." with concrete trigger phrases — the exact shape of the anchor-5 example. Not a 4 because the 'when' clause is fully explicit with enumerated triggers rather than merely present.

5 / 5

Trigger Term Quality

Covers natural phrases users would actually say across synonyms: "rebuttal", "reply to reviewers", "ICML rebuttal", "OpenReview response", "answer external reviews" — spanning generic and venue-specific variants. Not a 4 because the only misses are minor variants (e.g. 'author response'), which the anchor for 4 ('a few natural terms missing') treats as a larger gap.

5 / 5

Distinctiveness Conflict Risk

It occupies a clear niche (post-submission rebuttal for external peer reviews) with distinct triggers ('ICML rebuttal', 'OpenReview response') unlikely to fire for document-editing or code-review skills. Not a 4 because overlap risk with adjacent skills (e.g. generic 'review' skills) is minimal given the venue-specific vocabulary.

5 / 5

Total

20

/

20

Passed

Validation

87%

Checks the skill against the spec for correct structure and formatting. All validation checks must pass before discovery and implementation can be scored.

Validation — 14 / 16 Passed

Validation for skill structure

CriteriaDescriptionResult

allowed_tools_field

'allowed-tools' contains unusual tool name(s)

Warning

frontmatter_unknown_keys

Unknown frontmatter key(s) found; consider removing or moving to metadata

Warning

Total

14

/

16

Passed

Repository
wanshuiyin/Auto-claude-code-research-in-sleep
Reviewed

Table of Contents

Is this your skill?

If you maintain this skill, you can claim it as your own. Once claimed, you can manage eval scenarios, bundle related skills, attach documentation or rules, and ensure cross-agent compatibility.