CtrlK
BlogDocsLog inGet started
Tessl Logo

alfworld-open-receptacle

This skill opens a closed receptacle to access its contents. It should be triggered when an agent needs to interact with items inside a closed container (e.g., fridge, microwave, drawer). The skill takes a receptacle identifier as input, performs the open action, and outputs the observation of the interior, enabling subsequent item retrieval or placement.

64

Quality

80%

Does it follow best practices?

Run evals on this skill

Adds up to 20 points to the overall score

View guide
SecuritybySnyk

Passed

No findings from the security scan

Fix and improve this skill with Tessl

tessl review fix ./experiments/src/skills/alfworld/alfworld-open-receptacle/SKILL.md
SKILL.md
Quality
Evals
Security

Quality

Content

71%Weight 40%Scale 1-5

Reviews the quality of instructions and guidance provided to agents. Good implementation is clear, handles edge cases, and produces reliable results.

A concise, well-sectioned instruction skill with concrete commands, a sensible check-first workflow, and explicit success/failure outcomes. Two concrete defects hold it back: the 'Example from Trajectory' section is an empty stub, and the existing references/common_receptacles.md file — the natural home for the naming conventions the skill depends on — is never referenced from the body.

Suggestions

Fill in or remove the empty '## Example from Trajectory' section; a single concrete example (observation → 'open fridge 1' → resulting interior observation) would also raise actionability.

Link references/common_receptacles.md from the Input Format section (e.g., 'See [common_receptacles.md](references/common_receptacles.md) for the full list of receptacle names and their contents') so the existing bundle file is discoverable.

Add a one-line recovery step under Expected Outcomes: on 'Nothing happened', re-check the receptacle identifier against the observation and retry, closing the feedback loop.

DimensionReasoningScore

Conciseness

The body is lean — short Purpose, bulleted When to Use, and a four-step execution list with no explanation of concepts Claude already knows. Not a 5 because the trailing '## Example from Trajectory' header is empty dead weight: a token spend that earns nothing, matching anchor 4's 'minor instances that could be trimmed' rather than 'every token earns its place'.

4 / 5

Actionability

Concrete executable commands are given ('go to {recep}', 'open {recep}') plus a concrete failure signal ('"Nothing happened" indicates the action was invalid'). Not a 5 because the '## Example from Trajectory' section is empty, so there is no worked example showing an actual input, command, and resulting observation — a visible gap in copy-paste-ready coverage; clearly above anchor 3 since the guidance given is executable, not pseudocode.

4 / 5

Workflow Clarity

The four steps are clearly sequenced with a precondition check (verify the receptacle is 'closed') and expected success/failure outcomes that support error interpretation ('Nothing happened' → already open or wrong identifier). Not a 5 because there is no explicit feedback loop instructing what to do on failure (e.g., re-check the identifier and retry); above anchor 4's baseline is not justified without that recovery step, and it does not reach anchor 3 since checkpoints are present rather than missing.

4 / 5

Progressive Disclosure

Sections are well-organized for a sub-50-line skill, but the bundle contains references/common_receptacles.md — which documents exactly the naming conventions the Input Format section relies on — and the body never mentions or links it, so the reference is invisible to a reader. This matches anchor 3 ('references present but not clearly signaled'): structure exists, but an available reference file is completely unsignaled. Not anchor 4, which requires references to be 'mostly clear'.

3 / 5

Total

15

/

20

Passed

Description

80%Weight 40%Scale 1-5

Based on the skill's description, can an agent find and select it at the right time? Clear, specific descriptions lead to better discovery.

A strong description with an explicit what, an explicit when-to-use clause containing concrete container examples, and a clear input/output contract in third-person voice. The main weakness is trigger breadth: 'interact with items inside a closed container' overlaps with retrieval and placement skills rather than being scoped to the receptacle being closed.

Suggestions

Narrow the trigger condition to the receptacle's state (e.g., 'Use when an observation reports a receptacle is closed, e.g. "The fridge 1 is closed", and its contents must be accessed') to reduce overlap with take/put item skills.

Add one or two more receptacle synonyms (cabinet, safe) to broaden natural trigger coverage without changing scope.

DimensionReasoningScore

Specificity

The description lists several concrete specifics — 'takes a receptacle identifier as input, performs the open action, and outputs the observation of the interior' — with a clear I/O contract. It is not a 5 because the skill really performs only one action (open), so coverage of capabilities is narrow rather than comprehensive; it is above a 3 because the input/action/output detail goes beyond naming 1-2 generic actions.

4 / 5

Completeness

Both halves are explicit: 'This skill opens a closed receptacle to access its contents' answers what, and 'It should be triggered when an agent needs to interact with items inside a closed container (e.g., fridge, microwave, drawer)' answers when with concrete trigger phrasing. This matches the anchor-5 pattern of an explicit what plus an explicit when with concrete trigger conditions.

5 / 5

Trigger Term Quality

Natural trigger vocabulary is present: 'closed receptacle', 'items inside a closed container', and concrete examples 'fridge, microwave, drawer' that an agent would actually see in observations. Not a 5 because common receptacle synonyms (cabinet, safe, drawerbath) and phrasing like 'the fridge is closed' are missing; clearly above a 3 since several natural terms beyond generic domain language are included.

4 / 5

Distinctiveness Conflict Risk

The skill has a niche (opening receptacles) but its trigger condition — 'needs to interact with items inside a closed container' — would also fire for item-retrieval and item-placement skills that involve the same containers, creating real overlap risk. It is more specific than anchor 2's 'very broad' but does not reach anchor 4's 'mostly distinct' because the trigger overlaps with closely related take/put skills.

3 / 5

Total

16

/

20

Passed

Validation

100%

Checks the skill against the spec for correct structure and formatting. All validation checks must pass before discovery and implementation can be scored.

Validation — 16 / 16 Passed

Validation for skill structure

No warnings or errors.

Repository
zjunlp/SkillNet
Reviewed

Table of Contents

Is this your skill?

If you maintain this skill, you can claim it as your own. Once claimed, you can manage eval scenarios, bundle related skills, attach documentation or rules, and ensure cross-agent compatibility.