CtrlK
BlogDocsLog inGet started
Tessl Logo

research-fallback

Handle web research failures by pivoting to internal knowledge for content generation

49

Quality

53%

Does it follow best practices?

Run evals on this skill

Adds up to 20 points to the overall score

View guide

SecuritybySnyk

Passed

No findings from the security scan

Fix and improve this skill with Tessl

tessl review fix ./benchmarks/gdpval/skills/research-fallback/SKILL.md
SKILL.md
Quality
Evals
Security

Quality

Content

57%

Reviews the quality of instructions and guidance provided to agents. Good implementation is clear, handles edge cases, and produces reliable results.

The content is a well-organized, self-contained workflow with clear sequencing and an explicit pivot checkpoint, but it carries some redundancy across its example and best-practices sections and several steps stay at an abstract level rather than fully actionable.

Suggestions

Trim redundancy by merging 'Best Practices' into the workflow steps or removing the restated 'Example Application', since both restate the numbered process.

Make the abstract steps more concrete, e.g. replace 'Assess internal knowledge' with 'List the specific facts, frameworks, and dates you can state confidently from training' and 'Identify gaps' with 'Mark any claim dependent on current sources with a verification note'.

Add an explicit feedback/checklist checkpoint, such as a short checklist confirming disclaimers and uncertainty flags are present before finalizing the deliverable, to strengthen the workflow-clarity feedback loop.

DimensionReasoningScore

Conciseness

The body does not explain concepts Claude already knows (no definition of web research or tools), but the 'Best Practices', 'Example Decision Flow', and 'Example Application' sections partially restate the numbered workflow steps, adding length beyond what is strictly necessary. This matches the score-2 'mostly efficient but could be tightened' anchor rather than the lean, every-token-earns-its-place score of 3.

2 / 3

Actionability

There are concrete directives such as 'After 1-2 attempts, proceed to Step 3' and 'Include a note that external verification is recommended' and 'Avoid making specific claims about very recent events', but several steps remain abstract ('Assess internal knowledge', 'Identify gaps'). As an instruction-only skill the absence of code is acceptable, yet the mix of concrete and abstract guidance places it at score 2 rather than the fully-copy-paste-ready score of 3.

2 / 3

Workflow Clarity

The five steps are clearly sequenced with an explicit pivot decision checkpoint and retry bound ('after 1-2 attempts'), but there is no validate-then-fix-then-retry feedback loop of the kind the score-3 anchor requires. The sequence is present with an implicit checkpoint, matching score 2 rather than the full feedback-loop score of 3.

2 / 3

Progressive Disclosure

No bundle files exist and the skill is a single self-contained SKILL.md with well-organized, clearly labeled sections (When to Use, Workflow Steps, Best Practices, Example, When NOT to Use) and no nested references or content that should be split out. This matches the score-3 'clear overview, well-organized, easy navigation' anchor for a self-contained conceptual skill.

3 / 3

Total

9

/

12

Passed

Description

50%

Based on the skill's description, can an agent find and select it at the right time? Clear, specific descriptions lead to better discovery.

The description is specific and in correct third person, naming a clear action and trigger condition, but it lacks an explicit 'Use when...' clause and only covers a single action path, which limits it across the specificity, trigger, and completeness dimensions.

Suggestions

Add an explicit 'Use when...' clause, e.g. 'Use when web search or webpage reading tools fail after retries and the task can proceed from internal knowledge.'

Broaden trigger terms to natural phrasings users actually say, such as 'web search failed', 'tool unavailable', or 'research tools not working'.

List multiple concrete actions (e.g. 'Detect tool failures, pivot to internal knowledge, and generate content with verification disclaimers') to raise specificity.

DimensionReasoningScore

Specificity

The phrase 'Handle web research failures by pivoting to internal knowledge for content generation' names a concrete domain (web research failures) and a single action path (pivot to internal knowledge), but it does not list multiple distinct concrete actions as required for a score of 3. It is not vague like the score-1 anchor 'Helps with documents', so it clears score 2 but not score 3.

2 / 3

Completeness

It states the 'what' (pivot to internal knowledge for content generation) and implies the 'when' (web research failures), but there is no explicit 'Use when...' trigger clause. Per the judging guidelines, a missing explicit trigger clause caps completeness at 2, which is where this lands rather than the explicit-both-what-and-when score of 3.

2 / 3

Trigger Term Quality

Terms like 'web research failures', 'internal knowledge', and 'content generation' are relevant, but a user would more naturally say 'web search failed', 'tool unavailable', or 'research not working'. It includes some relevant keywords but is missing common natural variations, matching the score-2 anchor rather than the broad natural-term coverage of score 3.

2 / 3

Distinctiveness Conflict Risk

The niche (research-failure fallback) is fairly specific and unlikely to trigger for unrelated skills, but it could still overlap with general research or content-generation skills, matching the score-2 'somewhat specific but could still overlap' anchor rather than the clearly-distinct score-3 anchor.

2 / 3

Total

8

/

12

Passed

Validation

100%

Checks the skill against the spec for correct structure and formatting. All validation checks must pass before discovery and implementation can be scored.

Validation16 / 16 Passed

Validation for skill structure

No warnings or errors.

Repository
HKUDS/OpenSpace
Reviewed

Table of Contents

Is this your skill?

If you maintain this skill, you can claim it as your own. Once claimed, you can manage eval scenarios, bundle related skills, attach documentation or rules, and ensure cross-agent compatibility.