CtrlK
BlogDocsLog inGet started
Tessl Logo

catalysis-prior-art-and-benchmarking

Use this skill when a catalysis workflow needs literature-grounded benchmark papers, prior-art mapping, representative method conventions, or evidence-backed comparisons.

64

Quality

76%

Does it follow best practices?

Run evals on this skill

Adds up to 20 points to the overall score

View guide

SecuritybySnyk

Low

Low-risk findings worth noting

Fix and improve this skill with Tessl

tessl review fix ./skills/catalysis-prior-art-and-benchmarking/SKILL.md
SKILL.md
Quality
Evals
Security

Quality

Content

78%Weight 40%Scale 1-5

Reviews the quality of instructions and guidance provided to agents. Good implementation is clear, handles edge cases, and produces reliable results.

A concise, well-structured instruction-only skill with a concrete tool-depth decision table and clear method-critical defaults and output contract. Its weaknesses are mild: some duplicated web-vs-tool guidance across sections, abstractly stated workflow steps, and no example queries or sample output pack to anchor execution.

Suggestions

Consolidate the web-search-vs-run_literature_research guidance into one place (Quick Start or Workflow step 2) to remove the duplicated preference statement.

Add one short example literature question and a skeletal example of the returned literature pack (papers, convention summary, citations, open questions) to make the Output Contract concrete.

Add a brief self-check step to the Workflow (e.g. verify the pack covers benchmark conventions and flags disagreements) so the sequence ends with an explicit checkpoint.

DimensionReasoningScore

Conciseness

The body is lean (~40 lines), assumes domain competence, and contains no explanation of concepts Claude already knows. Minor trimmable redundancy remains — the web-vs-tool preference appears in both Quick Start ('online/web search may be enough') and Workflow step 2 ('Prefer plain web/online exploration'), and the depth=quick guidance is near-duplicated between Quick Start and step 2's 'use run_literature_research when you need paper-level citations'.

4 / 5

Actionability

The Quick Start gives concrete, executable guidance — 'call run_literature_research with depth=quick' and a depth-to-purpose mapping (quick/standard/focused/deep_report) — which for an instruction-only skill is concrete command-level guidance. It misses a 5 because there are no example literature queries or a concrete example of the output 'literature pack' format, leaving key execution details implicit.

4 / 5

Workflow Clarity

The three workflow steps are numbered and clearly sequenced (frame the question → align with stage → extract conventions), and the depth-selection decision path in Quick Start is unambiguous. This is a non-destructive, non-batch task so no validation checkpoint cap applies; it falls short of 5 because steps are stated abstractly ('Frame the query around the actual scientific need') with no worked example or checkpoint showing how to verify the pack meets the Output Contract.

4 / 5

Progressive Disclosure

Per the simple-skill exception (under 50 lines, no external reference files needed — no references/, scripts/, or assets/ directories exist), the well-organized section structure (Overview, Quick Start, Suggested tools, Workflow, Method-critical defaults, Output Contract, References) is sufficient. All cross-references are to sibling skills, not to missing or nested files, so navigation is trivial.

5 / 5

Total

17

/

20

Passed

Description

75%Weight 40%Scale 1-5

Based on the skill's description, can an agent find and select it at the right time? Clear, specific descriptions lead to better discovery.

A solid, niche-specific description with an explicit 'Use this skill when...' trigger and a concrete list of deliverables. Its main weaknesses are that the 'what' is expressed as needs rather than actions, and natural trigger-term coverage lacks synonyms a user might actually say.

Suggestions

State the 'what' with an action verb up front, e.g. 'Finds and summarizes representative catalysis papers, benchmark conventions, and prior art... Use when a catalysis workflow needs...'.

Add natural synonyms users might say, such as 'papers', 'literature search', 'reference systems', or 'comparable catalysts', to broaden trigger coverage.

Sharpen distinctiveness from the sibling literature-grounding skill by naming the catalysis-specific boundary (benchmark conventions, method defaults) in the description itself.

DimensionReasoningScore

Specificity

The description lists several concrete deliverables — 'literature-grounded benchmark papers, prior-art mapping, representative method conventions, or evidence-backed comparisons' — which are specific items within the catalysis domain, matching the 'several specific actions; minor gaps' anchor. It stops short of a 5 because these are outcomes the workflow 'needs' rather than clearly enumerated actions the skill performs.

4 / 5

Completeness

An explicit 'when' clause is present ('Use this skill when a catalysis workflow needs...') and the 'what' is conveyed through the list of deliverables, fitting the 'both what and when; one could be more explicit' anchor. The 'what' is implied rather than stated as an action verb (e.g. 'Finds and summarizes...'), keeping it below 5.

4 / 5

Trigger Term Quality

Natural domain terms a catalysis user would say are present: 'catalysis', 'benchmark papers', 'literature', 'prior-art mapping'. It is not a 5 because common synonyms and variations (e.g. 'papers', 'references', 'comparison systems', 'method conventions' phrasing users might actually type) have thin coverage, and some phrasing ('evidence-backed comparisons') is more crafted than natural.

4 / 5

Distinctiveness Conflict Risk

The catalysis literature-benchmarking niche is well defined and distinct from generic skills, but it sits adjacent to a general literature-grounding skill (the body even says 'Pair with literature-grounding'), creating minor overlap risk with a closely related skill — exactly the anchor-4 condition rather than the minimal-conflict anchor 5.

4 / 5

Total

16

/

20

Passed

Validation

93%

Checks the skill against the spec for correct structure and formatting. All validation checks must pass before discovery and implementation can be scored.

Validation — 15 / 16 Passed

Validation for skill structure

CriteriaDescriptionResult

metadata_version

'metadata.version' is missing

Warning

Total

15

/

16

Passed

Repository
fernandezbaptiste/CatMaster
Reviewed

Table of Contents

Is this your skill?

If you maintain this skill, you can claim it as your own. Once claimed, you can manage eval scenarios, bundle related skills, attach documentation or rules, and ensure cross-agent compatibility.