CtrlK
BlogDocsLog inGet started
Tessl Logo

serving-runtime-config

Configure custom ServingRuntime CRs on OpenShift AI for model serving frameworks not covered by built-in runtimes. Use when: - "Create a custom serving runtime" - "I need a runtime for ONNX / Triton / custom framework" - "Customize vLLM runtime parameters" - "What serving runtimes are available?" - "Add a custom container image for model serving" Handles listing existing runtimes, creating new ServingRuntime CRs, and validating compatibility with target models. NOT for deploying models (use /model-deploy after runtime is configured). NOT for NIM platform setup (use /nim-setup).

74

Quality

92%

Does it follow best practices?

Run evals on this skill

Adds up to 20 points to the overall score

View guide

SecuritybySnyk

Low

Low-risk findings worth noting

SKILL.md
Quality
Evals
Security

Quality

Content

88%Weight 40%Scale 1-5

Reviews the quality of instructions and guidance provided to agents. Good implementation is clear, handles edge cases, and produces reliable results.

A well-engineered operational skill: concrete MCP tooling with fallbacks, a complete YAML template, explicit HITL checkpoints, and validation. The main weaknesses are minor verbosity (duplicated fallback block, scripted announcement lines) and a bundle where most referenced files are empty pointer stubs, which undermines the otherwise good progressive-disclosure design.

Suggestions

State the rhoai-unavailable fallback for list_serving_runtimes once (e.g., in references/skill-conventions.md) instead of repeating the identical paragraph verbatim in Step 2 and Step 6.

Remove the scripted 'Output to user: "I consulted [supported-runtimes.md]..."' lines — they consume tokens without adding capability.

Fix the references/ bundle: common-issues.md, live-doc-lookup.md, openshift-fallback-templates.md, and skill-conventions.md currently contain only a bare path string rather than their referenced content, so several workflow steps (fallback templates, common-issues lookup, HITL conventions) resolve to dead ends.

DimensionReasoningScore

Conciseness

The body is almost entirely operational (tool names, parameters, fallback queries, error messages) and assumes Claude's knowledge of Kubernetes/OpenShift, with no concept explanations. Minor trim targets remain: the identical rhoai-fallback paragraph is repeated verbatim in Step 2 and Step 6, and scripted 'Output to user' lines like "I consulted [supported-runtimes.md]..." pad the workflow without adding capability. Not a 3 because no section is genuinely over-explained; not a 5 because the duplication and announcement scripting are noticeable.

4 / 5

Actionability

The skill provides a complete copy-paste ServingRuntime YAML manifest with all spec fields (labels, supportedModelFormats, secretKeyRef env, GPU resources), exact MCP tool names with REQUIRED parameter lists, concrete fallback queries (e.g., `apiVersion: serving.kserve.io/v1alpha1`, `labelSelector: opendatahub.io/dashboard=true`), an example template name ("vllm-cuda-runtime-template"), and exact user-facing error strings. Placeholders like [runtime-name] are justified by the preceding parameter-collection table with defaults, matching the anchor for fully executable guidance covering the common cases.

5 / 5

Workflow Clarity

Six explicitly sequenced steps each name the MCP tool, parameters, error handling, and end with WAIT-for-user checkpoints; Step 4 has a yes/no/modify feedback loop, Step 6 validates the created runtime, and the closing HITL section recapitulates every checkpoint including "NEVER overwrite an existing ServingRuntime without user confirmation". This matches the anchor of clear sequence with explicit validation, feedback loops, and checklists — not a 4, since no validation checkpoint is missing or merely implicit.

5 / 5

Progressive Disclosure

The body is a clear overview with well-signaled, one-level-deep references (supported-runtimes.md, live-doc-lookup.md, openshift-fallback-templates.md, common-issues.md, skill-conventions.md), each annotated with its purpose. However, scoring against the actual bundle shows only supported-runtimes.md has real content; the other referenced files are stubs containing just a path string, so navigation dead-ends, and references/known-model-profiles.md is orphaned (unreferenced). Good structure with organization gaps, matching anchor 4 rather than anchor 5's 'easy navigation'.

4 / 5

Total

18

/

20

Passed

Description

96%Weight 40%Scale 1-5

Based on the skill's description, can an agent find and select it at the right time? Clear, specific descriptions lead to better discovery.

An exemplary description: concrete third-person capability statements, five natural trigger phrases with framework synonyms, explicit what/when coverage, and clear negative boundaries against adjacent skills. The only weakness is slight overlap risk with /model-deploy on the runtime-listing trigger.

DimensionReasoningScore

Specificity

Multiple specific concrete actions in third person: "Configure custom ServingRuntime CRs on OpenShift AI", "listing existing runtimes, creating new ServingRuntime CRs, and validating compatibility with target models". Coverage spans the full scope of the skill with no filler. Matches the anchor for multiple specific concrete actions with comprehensive coverage; a 4 would require a noticeable coverage gap, which is absent.

5 / 5

Completeness

Explicitly answers 'what' ("Configure custom ServingRuntime CRs... Handles listing existing runtimes, creating new ServingRuntime CRs, and validating compatibility") and 'when' (a concrete "Use when:" list of five trigger phrases), plus negative boundaries ("NOT for deploying models (use /model-deploy...)", "NOT for NIM platform setup"). This is a clear match to the anchor requiring both what and when with concrete trigger phrases.

5 / 5

Trigger Term Quality

The 'Use when' list gives five natural user phrasings — "Create a custom serving runtime", "I need a runtime for ONNX / Triton / custom framework", "Customize vLLM runtime parameters", "What serving runtimes are available?", "Add a custom container image for model serving" — including framework synonyms (ONNX, Triton, vLLM) and both imperative and question forms. This is comprehensive natural-term coverage for the domain (the domain has no file extensions to include), matching anchor 5 rather than anchor 4's 'a few natural terms missing'.

5 / 5

Distinctiveness Conflict Risk

The niche is clear (custom ServingRuntime configuration) and the NOT-for lines explicitly demarcate /model-deploy and /nim-setup, but the trigger "What serving runtimes are available?" would also naturally fire when a user is exploring runtimes to deploy a model — /model-deploy's territory — giving minor overlap risk with a closely related sibling skill. Matches 'mostly distinct; minor overlap risk with closely related skills' (4) rather than anchor 5's 'minimal conflict risk'.

4 / 5

Total

19

/

20

Passed

Validation

93%

Checks the skill against the spec for correct structure and formatting. All validation checks must pass before discovery and implementation can be scored.

Validation — 15 / 16 Passed

Validation for skill structure

CriteriaDescriptionResult

frontmatter_unknown_keys

Unknown frontmatter key(s) found; consider removing or moving to metadata

Warning

Total

15

/

16

Passed

Repository
RHEcosystemAppEng/agentic-plugins
Reviewed

Table of Contents

Is this your skill?

If you maintain this skill, you can claim it as your own. Once claimed, you can manage eval scenarios, bundle related skills, attach documentation or rules, and ensure cross-agent compatibility.