CtrlK
BlogDocsLog inGet started
Tessl Logo

regulatory-fallback-research

Handle tool failures when researching regulatory/government content by using fallback methods and domain knowledge

50

Quality

63%

Does it follow best practices?

Run evals on this skill

Adds up to 20 points to the overall score

View guide
SecuritybySnyk

Low

Low-risk findings worth noting

Fix and improve this skill with Tessl

tessl review fix ./benchmarks/gdpval/skills/regulatory-fallback-research/SKILL.md
SKILL.md
Quality
Evals
Security

Quality

Content

60%Weight 40%Scale 1-5

Reviews the quality of instructions and guidance provided to agents. Good implementation is clear, handles edge cases, and produces reliable results.

The body is a well-structured, clearly sequenced fallback workflow with a useful failure-mapping table, but its 'code examples' are illustrative pseudocode rather than executable commands and some sections are padded.

Suggestions

Replace pseudocode blocks with concrete, copy-paste-ready invocations (real shell_agent command syntax and actual function calls or templates) instead of 'create_pharmacy_checklist(...)' placeholders.

Tighten the Overview and Best Practices by removing explanatory padding about why government sites fail; assume Claude's competence.

Add an explicit validate/retry checkpoint in Step 1 (e.g. 'After N retries return unknown error, document the failure and proceed to Step 2') to make the feedback loop explicit.

DimensionReasoningScore

Conciseness

Mostly efficient procedural prose, but the Overview and Best Practices contain explanatory padding, and several fenced 'code' blocks are prose placeholders rather than real code, which could be tightened.

3 / 5

Actionability

Provides a clear step sequence and example blocks, but the examples are pseudocode (e.g. 'shell_agent task: "Research [topic]..."', 'create_pharmacy_checklist(...)') rather than executable commands, with placeholder tokens like [topic].

3 / 5

Workflow Clarity

A clearly sequenced 4-step procedure with expected outcomes and a failure-to-fallback mapping table; checkpoints are mostly present though some (e.g. when to stop retrying) are implicit rather than explicit validate-fix-retry loops.

4 / 5

Progressive Disclosure

Well-organized single-file skill with clear section headers and no nested references; minor gap is that ~140 lines of domain-knowledge lists and examples are inlined rather than split into reference files.

4 / 5

Total

14

/

20

Passed

Description

50%Weight 40%Scale 1-5

Based on the skill's description, can an agent find and select it at the right time? Clear, specific descriptions lead to better discovery.

The description clearly states the domain and general approach but lacks an explicit 'Use when...' trigger clause and uses generic action verbs, leaving it at the midpoint across all dimensions.

Suggestions

Add an explicit trigger clause, e.g. 'Use when primary regulatory sources (FDA, CMS, state boards) return errors or fail repeatedly'.

Replace generic verbs ('Handle', 'using fallback methods') with concrete actions like 'retry via shell_agent, then synthesize compliance deliverables from domain knowledge'.

Include natural user phrasings such as 'government site down', 'regulatory lookup failed', or 'compliance checklist from domain knowledge'.

DimensionReasoningScore

Specificity

Names the domain ('regulatory/government content') and 1-2 actions ('Handle tool failures', 'using fallback methods and domain knowledge') but the actions are generic and coverage is not comprehensive.

3 / 5

Completeness

Has a clear 'what' but only a weakly implied 'when' (the phrase 'when researching' is embedded rather than an explicit 'Use when...' clause), so completeness is capped at 3 per the rubric guideline.

3 / 5

Trigger Term Quality

Contains relevant keywords ('tool failures', 'regulatory/government content', 'fallback methods') but misses common natural variations a user would actually say (e.g. 'FDA site down', 'CMS lookup failed').

3 / 5

Distinctiveness Conflict Risk

The regulatory-research + tool-failure-fallback niche is somewhat specific, but the 'handle tool failures using fallback methods' framing could overlap with general web-research skills.

3 / 5

Total

12

/

20

Passed

Validation

100%

Checks the skill against the spec for correct structure and formatting. All validation checks must pass before discovery and implementation can be scored.

Validation — 16 / 16 Passed

Validation for skill structure

No warnings or errors.

Repository
HKUDS/OpenSpace
Reviewed

Table of Contents

Is this your skill?

If you maintain this skill, you can claim it as your own. Once claimed, you can manage eval scenarios, bundle related skills, attach documentation or rules, and ensure cross-agent compatibility.