CtrlK
BlogDocsLog inGet started
Tessl Logo

scienceworld-conditional-placer

Places an object into one of several designated containers based on a measured condition, such as a temperature threshold. Use this skill when you have completed a measurement or assessment and the task requires sorting or storing the object into one of multiple containers according to a rule (e.g., "if temperature > X, place in container A; otherwise container B").

67

Quality

84%

Does it follow best practices?

Run evals on this skill

Adds up to 20 points to the overall score

View guide
SecuritybySnyk

Passed

No findings from the security scan

SKILL.md
Quality
Evals
Security

Quality

Content

78%Weight 40%Scale 1-5

Reviews the quality of instructions and guidance provided to agents. Good implementation is clear, handles edge cases, and produces reliable results.

A concise, highly actionable skill body with an excellent worked example and clear step sequencing. The main gaps are the orphaned references/action_guide.md, which is never linked from the body (leaving its additional actions undiscoverable and the Key Actions table duplicating part of it), and the absence of a small validation checkpoint on the measurement result.

Suggestions

Link references/action_guide.md from the body (e.g., under Key Actions: 'Full action reference: see references/action_guide.md'), so its additional actions like 'examine OBJ' and 'look at OBJ' are discoverable.

De-duplicate the Key Actions table against the Core Workflow steps, or replace it with a pointer to the action guide to save tokens.

Add a brief checkpoint before placement, e.g., 'If the measurement returns no value, verify both objects are in inventory and re-run use OBJ on OBJ'.

DimensionReasoningScore

Conciseness

The body is lean and assumes competence (no concept explanations), but the Key Actions table ('teleport to LOC', 'pick up OBJ', 'use OBJ on OBJ', 'move OBJ to OBJ') restates commands already given step-by-step in Core Workflow, which is a minor instance of duplication that could be trimmed. Not 5 due to this redundancy; well above the 'noticeably verbose' level.

4 / 5

Actionability

Guidance is fully executable: exact environment commands in every step, and a complete worked example with real values ('use thermometer on metal fork — reads 72 degrees', '72 > 50, so: move metal fork to orange box'). This matches the anchor for copy-paste-ready commands with a specific example covering the common case.

5 / 5

Workflow Clarity

The five-step Core Workflow is clearly sequenced with concrete commands, and the Important Notes flag a real pitfall (containers are pre-opened; do not use open/close). It falls short of 5 because there is no checkpoint confirming the measurement returned a usable value before evaluating the condition, but the operation is neither destructive nor batch, so no cap applies.

4 / 5

Progressive Disclosure

The body is well organized into Purpose / Core Workflow / Key Actions / Example / Important Notes, but the bundle contains references/action_guide.md which is never mentioned or linked anywhere in the body — its additional actions (e.g., 'examine OBJ', 'look at OBJ') are undiscoverable from SKILL.md, and the inline Key Actions table duplicates content that also lives in that reference. Per the guideline to score against the actual bundle structure, this fits 'references present but not clearly signaled; content that should be separate is inline' rather than the well-signaled level.

3 / 5

Total

16

/

20

Passed

Description

83%Weight 40%Scale 1-5

Based on the skill's description, can an agent find and select it at the right time? Clear, specific descriptions lead to better discovery.

A strong description that explicitly states what the skill does and when to use it, with concrete actions and a specific trigger rule example. Third-person voice is used correctly and there is no padding. The only weaknesses are minor: trigger synonyms are somewhat narrow and the condition types are exemplified only via temperature.

DimensionReasoningScore

Specificity

The description lists several concrete actions — placing an object into designated containers, evaluating a measured condition ('such as a temperature threshold'), and 'sorting or storing the object into one of multiple containers according to a rule' — with a concrete rule example. It falls short of 5 because coverage of condition types is exemplified only by temperature, but exceeds 3 because more than 1-2 actions are concretely specified.

4 / 5

Completeness

Both questions are answered explicitly: what ('Places an object into one of several designated containers based on a measured condition') and when ('Use this skill when you have completed a measurement or assessment and the task requires sorting or storing the object...'), including a concrete trigger rule example. This matches the anchor for clearly and explicitly answering both with concrete trigger phrases.

5 / 5

Trigger Term Quality

Natural phrases a user would say are present: 'measurement', 'temperature threshold', 'sorting or storing', 'place in container'. A few natural variations (e.g., 'categorize', 'put away', other measurement types like weight or conductivity) are missing, matching the 'good keyword coverage; a few natural terms missing' anchor rather than comprehensive coverage.

4 / 5

Distinctiveness Conflict Risk

The measured-condition rule ('if temperature > X, place in container A; otherwise container B') carves out a clear niche distinct from generic placement skills, but the 'sorting or storing' phrasing leaves minor overlap risk with generic object-sorting skills, fitting 'mostly distinct; minor overlap risk'.

4 / 5

Total

17

/

20

Passed

Validation

100%

Checks the skill against the spec for correct structure and formatting. All validation checks must pass before discovery and implementation can be scored.

Validation — 16 / 16 Passed

Validation for skill structure

No warnings or errors.

Repository
zjunlp/SkillNet
Reviewed

Table of Contents

Is this your skill?

If you maintain this skill, you can claim it as your own. Once claimed, you can manage eval scenarios, bundle related skills, attach documentation or rules, and ensure cross-agent compatibility.