CtrlK
BlogDocsLog inGet started
Tessl Logo

opik-integrations

Build, update, test, and document Opik SDK integrations (Python & TypeScript). Use when adding a new framework/provider integration under sdks/python/src/opik/integrations or sdks/typescript/src/opik/integrations, updating an existing one, or verifying that an integration logs traces correctly.

67

Quality

84%

Does it follow best practices?

Run evals on this skill

Adds up to 20 points to the overall score

View guide
SecuritybySnyk

Passed

No findings from the security scan

SKILL.md
Quality
Evals
Security

Quality

Content

78%Weight 40%Scale 1-5

Reviews the quality of instructions and guidance provided to agents. Good implementation is clear, handles edge cases, and produces reliable results.

A well-written overview with an excellent disambiguation note, a genuinely useful clone-sibling decision table, and a disciplined phased workflow. Its one real defect is that the progressive-disclosure chain dangles: the referenced workflow.md, python.md, and typescript.md are not present in the bundle, so the skill's depth is inaccessible.

Suggestions

Ship the referenced bundle files (workflow.md, python.md, typescript.md) alongside SKILL.md, or inline the minimal critical content (phase checklist, test command, MCP verification step) so the skill works standalone.

Add an explicit feedback loop in the body — e.g., 'if MCP verification shows a wrong trace/span tree, fix the capture logic and re-verify' — rather than relying on workflow.md for error recovery.

Include one concrete example of the golden rule in action (target → chosen sibling → what was adapted) so the clone-first instruction is executable without the language reference files.

DimensionReasoningScore

Conciseness

The body is lean throughout: it assumes Claude knows what tracing, OTel, and SDKs are and spends tokens only on non-obvious routing decisions — the "Do not confuse this with the user-facing instrument/opik skills" note, the clone-the-closest-sibling decision table, and "OpenTelemetry is backend-first" guidance. Every section earns its place; no padding.

5 / 5

Actionability

Highly actionable instruction-only content: a decision table mapping target shape to concrete patterns and clone sources ("openai/ · opik-openai", "langchain/ · opik-langchain"), a specific pre-check ("check whether track_openai(..., provider=...) already covers the need"), and concrete verification via the Opik MCP ("read/list the trace & spans"). Not 5 because the executable specifics of testing, MCP verification, and documentation wiring are delegated to referenced files rather than present in the body.

4 / 5

Workflow Clarity

A clearly sequenced 0-8 phase list with named checkpoints — "Verify the logged data through the Opik MCP", "Test with the language's integration-test harness", "Report — a high-level summary... with evidence" — plus up-front input collection in the questionnaire. Not 5 because the explicit validate-then-fix-then-retry feedback loop and the report template are only pointed at in workflow.md rather than stated in the body; not 3 because validation checkpoints are explicit and well-placed.

4 / 5

Progressive Disclosure

The in-body structure is excellent — clear overview, well-labeled one-level-deep references ([workflow.md], [python.md], [typescript.md]) with accurate content hints — but scored against the actual bundle: none of the three referenced files exist in the skill directory, so the core detail (the full playbook, test harnesses, OTel sections) is unreachable and navigation is broken. Good signaling cannot compensate for missing targets, placing this below the 4 anchor's 'references mostly clear'.

3 / 5

Total

16

/

20

Passed

Description

83%Weight 40%Scale 1-5

Based on the skill's description, can an agent find and select it at the right time? Clear, specific descriptions lead to better discovery.

A strong description: explicit what and when, concrete trigger phrases, and repo-path anchoring that makes routing unambiguous for SDK contributors. The only gaps are missing synonyms (tracing/instrumentation) and no in-description disambiguation from sibling Opik integration skills.

DimensionReasoningScore

Specificity

"Build, update, test, and document Opik SDK integrations (Python & TypeScript)" lists four concrete actions, and the trigger clause anchors them to real paths ("sdks/python/src/opik/integrations or sdks/typescript/src/opik/integrations"). Not a 5 because the actions are broad verbs — the description never names finer-grained capabilities like wiring trace/span verification, pattern selection, or documentation routing — but well above 3's '1-2 concrete actions'.

4 / 5

Completeness

Explicitly answers both: what — "Build, update, test, and document Opik SDK integrations (Python & TypeScript)" — and when — "Use when adding a new framework/provider integration under [paths], updating an existing one, or verifying that an integration logs traces correctly", with concrete trigger phrases and repo locations.

5 / 5

Trigger Term Quality

Good natural phrasing users would actually say: "adding a new framework/provider integration", "updating an existing one", "verifying that an integration logs traces correctly", plus concrete path triggers. Not 5 because common synonyms like "tracing", "instrumentation", "SDK contributor", or naming example frameworks (LangChain, OpenAI) are absent from the description itself.

4 / 5

Distinctiveness Conflict Risk

The path-scoped niche ("under sdks/python/src/opik/integrations or sdks/typescript/src/opik/integrations") makes it highly distinct from generic skills. Not 5 because the description does not itself disambiguate from the closely related external-integration/instrumentation skills (that routing lives only in the body), leaving minor overlap risk for requests like 'add an Opik integration' without a stated location.

4 / 5

Total

17

/

20

Passed

Validation

93%

Checks the skill against the spec for correct structure and formatting. All validation checks must pass before discovery and implementation can be scored.

Validation — 15 / 16 Passed

Validation for skill structure

CriteriaDescriptionResult

relative_links

Relative link issues: 5 missing

Warning

Total

15

/

16

Passed

Repository
comet-ml/opik
Reviewed

Table of Contents

Is this your skill?

If you maintain this skill, you can claim it as your own. Once claimed, you can manage eval scenarios, bundle related skills, attach documentation or rules, and ensure cross-agent compatibility.