Install, configure, verify, and improve vLLM Semantic Router through its CLI and Router API. Use for deployment, recipe tuning, and single-model/MoM evaluation, with Dashboard verification when requested.
The canonical home for this skill is vllm-sr-agent-operations in vllm-project/semantic-router
Work against the user's selected stack and objective. Inspect the installed CLI, running configuration and available backends before choosing an approach. Preserve unrelated workloads, credentials and the user's existing authorization.
Use vllm-sr --help, command-specific help and vllm-sr config schema.
For a running Router, GET /api/v1 advertises its operations and schemas.
Management origin, inference listener and public model entrypoint are separate;
discover them instead of assuming default ports or a recipe name.
Read only the reference needed for the task:
| Task | Reference |
|---|---|
| Install, select hardware/runtime, isolate a stack, open Dashboard access | Deployment |
| Change live config, activate a recipe, recover a revision | Configuration |
| Verify routing, tools, context boundaries or delivery | Route verification |
| Improve signal, decision or model-selection policy | Recipe tuning |
| Compare single models and MoM; run a measured optimization loop | sr-bench |
For installation or an authorized upgrade, use the published stable package unless the user selects another version or the dev channel:
curl -fsSL https://vllm-sr.ai/install.sh | \
bash -s -- --channel stable --mode cli --runtime skip --no-launch
export PATH="$HOME/.local/bin:$PATH"
vllm-sr --versionserve. For an existing
stack, derive changes from fresh config get, then validate, plan and apply.
Respect restart-required changes and verify the active revision afterward.For Dashboard work, exercise the corresponding user flow against the same stack and inspect the resulting artifacts. Leave the user with the active config or recipe, access details, evidence and material limitations. Keep secret values and private request content out of public artifacts.
c89d5d4
Canonical home
If you maintain this skill, you can claim it as your own. Once claimed, you can manage eval scenarios, bundle related skills, attach documentation or rules, and ensure cross-agent compatibility.