CtrlK
BlogDocsLog inGet started
Tessl Logo

alfworld-inventory-management

Use when the agent must collect and track multiple instances of the same object type in ALFWorld (e.g., "put two cellphone in bed"). This skill maintains a count of collected versus needed objects, guides systematic searching through receptacles, and ensures each found object is placed at the target before searching for the next.

73

Quality

92%

Does it follow best practices?

Run evals on this skill

Adds up to 20 points to the overall score

View guide
SecuritybySnyk

Passed

No findings from the security scan

SKILL.md
Quality
Evals
Security

Quality

Content

88%Weight 40%Scale 1-5

Reviews the quality of instructions and guidance provided to agents. Good implementation is clear, handles edge cases, and produces reliable results.

A highly actionable, well-sequenced skill body with executable command templates, a full worked example, and explicit validation/recovery loops for a batch collection task. The two minor weaknesses are the redundant 'When to Use' section and an orphaned bundle file (references/common_mistakes.md) that is never linked from the body.

Suggestions

Trim or merge the 'When to Use' section, which repeats the frontmatter description nearly verbatim, to save tokens.

Add a clearly signaled reference to the bundle file, e.g. a 'Common mistakes' section with 'See [common_mistakes.md](references/common_mistakes.md) for trajectory-derived failure patterns', so the reference is discoverable.

De-duplicate the search-priority and immediate-placement guidance between the body and common_mistakes.md so each lives in one place.

DimensionReasoningScore

Conciseness

The body is lean and assumes Claude's competence — no general concept explanations, just ALFWorld-specific procedure (decision tree, action templates, worked transcript). It is not a 5 because the 'When to Use' section largely restates the frontmatter description and could be trimmed. Not a 3 because there is no genuinely unnecessary explanation.

4 / 5

Actionability

Fully executable guidance: exact ALFWorld command templates (`take {object} from {current_receptacle}`, `go to {target_receptacle}`, `put {object} in/on {target_receptacle}`), a step-by-step decision tree, and a complete worked example transcript covering the common two-object case. Copy-paste ready and covers the typical scenario.

5 / 5

Workflow Clarity

Clear four-phase sequence with explicit validation checkpoints: counter initialization, per-placement `collected += 1` updates, a `collected == needed` completion gate, and feedback loops for error recovery (revisit searched receptacles, re-examine target on counter mismatch, retry on 'Nothing happened'). This matches the top anchor despite being a batch operation, because validation and recovery loops are present.

5 / 5

Progressive Disclosure

Sections are well-organized and the body is appropriately sized, but the bundle contains `references/common_mistakes.md` which is never referenced or linked from the body — a minor organization/discovery gap. It is not a 5 because an existing reference file goes unsignaled; not a 3 because the body structure itself is good and content placement is appropriate, fitting 'minor organization gaps'.

4 / 5

Total

18

/

20

Passed

Description

92%Weight 40%Scale 1-5

Based on the skill's description, can an agent find and select it at the right time? Clear, specific descriptions lead to better discovery.

A strong description: third-person voice, explicit 'Use when' trigger with a verbatim task example, and three concrete capability statements. The only minor gap is that a few natural trigger synonyms (quantity/count phrasings) are missing.

DimensionReasoningScore

Specificity

The description lists multiple concrete actions — "maintains a count of collected versus needed objects", "guides systematic searching through receptacles", "ensures each found object is placed at the target before searching for the next" — which fully cover this skill's scope. It is not a 4 because the action list is comprehensive for the domain rather than having minor coverage gaps.

5 / 5

Completeness

It explicitly answers 'when' with an explicit "Use when the agent must collect and track multiple instances..." clause plus a concrete trigger example, and answers 'what' with three concrete capabilities. Both are explicit with concrete trigger phrases, matching the top anchor exactly.

5 / 5

Trigger Term Quality

Strong natural triggers: the quoted task phrasing "put two cellphone in bed" mirrors exactly what a user/agent would see, plus "collect", "track", "multiple instances of the same object type", "receptacles". It is not a 5 because common variations like quantity words ("two of", "several", "count") or "find and place" are absent; not a 3 because keyword coverage clearly exceeds 'some relevant keywords'.

4 / 5

Distinctiveness Conflict Risk

"in ALFWorld" and the multi-instance collection trigger carve out a clear niche with a verbatim task-example trigger, making conflict with other skills minimal. It is not a 4 because no closely related skill domain overlaps with this trigger set.

5 / 5

Total

19

/

20

Passed

Validation

100%

Checks the skill against the spec for correct structure and formatting. All validation checks must pass before discovery and implementation can be scored.

Validation — 16 / 16 Passed

Validation for skill structure

No warnings or errors.

Repository
zjunlp/SkillNet
Reviewed

Table of Contents

Is this your skill?

If you maintain this skill, you can claim it as your own. Once claimed, you can manage eval scenarios, bundle related skills, attach documentation or rules, and ensure cross-agent compatibility.