CtrlK
BlogDocsLog inGet started
Tessl Logo

benchling-integration

Benchling R&D platform integration. Access registry (DNA, proteins), inventory, ELN entries, workflows via API, build Benchling Apps, query Data Warehouse, for lab data management automation.

56

Quality

65%

Does it follow best practices?

Run evals on this skill

Adds up to 20 points to the overall score

View guide

SecuritybySnyk

Passed

No findings from the security scan

Fix and improve this skill with Tessl

tessl review fix ./backend/cli/skills/biology/benchling-integration/SKILL.md
SKILL.md
Quality
Evals
Security

Quality

Content

61%Weight 40%Scale 1-5

Reviews the quality of instructions and guidance provided to agents. Good implementation is clear, handles edge cases, and produces reliable results.

A well-structured, code-heavy body with genuine one-level-deep references and mostly executable guidance. Weaknesses are verbosity from duplicated example sections, batch-operation workflows lacking validation checkpoints (capping workflow clarity at 3), and a scripts/ section referencing files that don't exist.

Suggestions

Add validation/verification steps to the batch workflows — e.g., check API responses or re-query created entities after the bulk FASTA import, and handle failures per-record rather than crashing mid-loop.

Remove or fix the '### scripts/' section: it describes example scripts that do not exist in the bundle, which misleads navigation.

Trim the 'Common Use Cases' section (or move it to a reference file) — its four long code examples duplicate patterns already shown in the capability sections, tightening the body considerably.

DimensionReasoningScore

Conciseness

The body is mostly efficient — code-forward with no basic-concept padding Claude already knows — but at ~470 lines it includes unnecessary bulk: the 'Common Use Cases' section (four long code examples) and repeated 'Key Operations' bullet lists duplicate coverage already given in the section examples and in references/. It is not a 2 because the padding is structural redundancy rather than explanatory fluff, and not a 4 because multiple sections could clearly be trimmed or moved to the reference files.

3 / 5

Actionability

Concrete, largely copy-paste-ready Python throughout — auth setup, entity create/update, generator pagination, the fields() helper, wait_for_task — matching anchor 4. It falls short of anchor 5 because the Data Warehouse section ('Connect using standard SQL clients with provided credentials') and the Events section offer no executable code or commands, and a few SDK calls (e.g., benchling.containers.transfer) cannot be verified as real API surface.

4 / 5

Workflow Clarity

Sequences are present (e.g., the four-step EventBridge integration pattern), but batch operations — the bulk FASTA import loop and the bulk workflow-task update loop — include no validation or verification steps (no error handling, no confirmation the created entities landed correctly). Per the rubric guideline, missing validation in batch workflows caps workflow_clarity at 3; it is not a 2 because steps are listed coherently rather than being poorly defined.

3 / 5

Progressive Disclosure

Good structure: three real, one-level-deep reference files (authentication.md, sdk_reference.md, api_endpoints.md) clearly signaled both inline ('refer to references/authentication.md') and in a Resources section, matching anchor 4. It is not a 5 because the scripts/ section points to a directory that does not exist in the bundle, and a meaningful amount of example-heavy content (e.g., the Common Use Cases code) arguably belongs in the reference files rather than inlined in SKILL.md.

4 / 5

Total

14

/

20

Passed

Description

70%Weight 40%Scale 1-5

Based on the skill's description, can an agent find and select it at the right time? Clear, specific descriptions lead to better discovery.

A specific, well-scoped description for a named platform with strong distinctiveness, undermined mainly by the absence of an explicit 'Use when...' trigger clause, which caps completeness at 3. Keyword coverage is good but misses natural synonyms like 'plasmids', 'sequences', and 'lab notebook'.

Suggestions

Append an explicit trigger clause such as: 'Use when working with Benchling, its Python SDK or REST API, or when syncing lab data (sequences, samples, notebook entries) to or from Benchling.'

Add natural synonyms users would say — e.g., 'plasmids', 'biological sequences', 'lab notebook entries', 'samples' — to improve trigger term coverage.

Drop the vague tail 'for lab data management automation' and replace it with the concrete trigger phrasing above.

DimensionReasoningScore

Specificity

The description lists several concrete actions — 'Access registry (DNA, proteins), inventory, ELN entries, workflows via API, build Benchling Apps, query Data Warehouse' — matching the 'several specific actions; minor gaps' anchor. It falls short of anchor 5 because the actions are compressed fragments rather than fully expressed capabilities and the tail 'for lab data management automation' is generic filler.

4 / 5

Completeness

The 'what' is clear and multi-action, but there is no 'Use when...' clause or equivalent explicit trigger guidance; 'for lab data management automation' only weakly implies when to use it. Per the rubric guideline, a missing 'Use when...' clause caps completeness at 3, and this is not a 2 because the 'what' is concrete rather than vague.

3 / 5

Trigger Term Quality

Good keyword coverage with domain terms a user might say (Benchling, registry, DNA, proteins, inventory, ELN, workflows, Data Warehouse), matching anchor 4. It is not a 5 because common natural synonyms users would actually say — 'plasmids', 'sequences', 'lab notebook', 'samples' — are absent, though it is clearly above the midpoint given the named platform acts as the dominant trigger.

4 / 5

Distinctiveness Conflict Risk

It occupies a clear niche — the named Benchling platform — with distinct domain triggers (registry entities, ELN, Data Warehouse, Benchling Apps) that no generic skill would claim. Conflict risk is minimal, matching anchor 5 exactly.

5 / 5

Total

16

/

20

Passed

Validation

87%

Checks the skill against the spec for correct structure and formatting. All validation checks must pass before discovery and implementation can be scored.

Validation — 14 / 16 Passed

Validation for skill structure

CriteriaDescriptionResult

metadata_version

'metadata.version' is missing

Warning

frontmatter_unknown_keys

Unknown frontmatter key(s) found; consider removing or moving to metadata

Warning

Total

14

/

16

Passed

Repository
synthetic-sciences/openscience
Reviewed

Table of Contents

Is this your skill?

If you maintain this skill, you can claim it as your own. Once claimed, you can manage eval scenarios, bundle related skills, attach documentation or rules, and ensure cross-agent compatibility.