CtrlK
BlogDocsLog inGet started
Tessl Logo

managing-devops-pipeline

通过 MCP 管理 BK-CI 流水线构建时使用,例如查询构建历史、获取启动参数、查看构建状态和在确认后触发构建。当用户要操作现有流水线而不是修改代码实现时优先使用。

68

Quality

85%

Does it follow best practices?

Run evals on this skill

Adds up to 20 points to the overall score

View guide
SecuritybySnyk

Passed

No findings from the security scan

SKILL.md
Quality
Evals
Security

Quality

Content

71%Weight 40%Scale 1-5

Reviews the quality of instructions and guidance provided to agents. Good implementation is clear, handles edge cases, and produces reliable results.

A well-structured, token-efficient operational guardrail skill with an explicit confirmation gate for side-effect operations. Its main weaknesses are that all referenced detail files are missing from the bundle (so the main file alone is not fully executable) and the absence of error-recovery guidance.

Suggestions

Ship the referenced files (reference/build-list.md, build-startinfo.md, build-status.md, build-start.md) or remove/fix the 延伸阅读 links — as delivered, every reference path dead-ends and the deferred tool-call details are unreachable.

Add a short error-recovery loop for the trigger workflow: what to do when the start API fails, how to verify the build actually started (e.g., re-check build status with the returned buildId), and when to surface errors to the user.

Consolidate the triple-stated confirmation rule into one authoritative statement (e.g., in 高信号规则) and keep the 关键陷阱 entries as one-line cross-references to trim repetition.

DimensionReasoningScore

Conciseness

The ~50-line body is lean and assumes Claude's competence — no explanation of what BK-CI or MCP is, no filler — but the user-confirmation-before-trigger rule is stated three times ("先向用户展示完整入参并获得确认", "所有有副作用的构建启动操作,都必须先获得用户明确确认", "没展示完整入参就直接触发构建"), which is trimmable repetition. It fits 'efficient; minor instances of over-explanation that could be trimmed' rather than score 5's 'every token earns its place'.

4 / 5

Actionability

As an instruction-only skill the guidance is concrete: identifier formats are specified ("projectId:项目英文名", "pipelineId:以 p- 开头", "buildId:以 b- 开头") and the trigger chain is sequenced (get params → show full inputs → confirm → trigger). It is not score 5 because no tool names, URL formats, or example payloads appear anywhere in the bundle — everything executable is deferred to reference files that are not present — leaving minor gaps such as how to recognize a pipeline detail-page URL.

4 / 5

Workflow Clarity

The core workflow is clearly sequenced with the critical validation checkpoint explicit: "先拿启动参数,再向用户展示完整入参并获得确认,最后才真正启动", plus a read-vs-trigger distinction and a mandatory identifier check. It is not score 5 because there is no error-recovery feedback loop (e.g., what to do when a trigger call fails or a build errors) and no post-trigger verification step, matching 'clear sequence with most checkpoints present; minor validation gaps'.

4 / 5

Progressive Disclosure

The 延伸阅读 section clearly signals one-level-deep, per-task references ("获取构建历史:reference/build-list.md" etc.), but none of the referenced files exist in the skill bundle — there is no reference/ or references/ directory at all — so navigation dead-ends. Per the guideline to score against the actual bundle structure, references that do not resolve keep this at 'some structure but could be better organized' rather than score 4's 'references mostly clear' with only minor gaps.

3 / 5

Total

15

/

20

Passed

Description

92%Weight 40%Scale 1-5

Based on the skill's description, can an agent find and select it at the right time? Clear, specific descriptions lead to better discovery.

A strong description: concrete action list, explicit use-when guidance with a boundary against adjacent skills, and natural trigger vocabulary in the target language. The only improvement space is adding a few synonyms (e.g., 蓝盾/执行记录) for broader trigger coverage.

DimensionReasoningScore

Specificity

The description lists four concrete actions — "查询构建历史" (query build history), "获取启动参数" (get startup parameters), "查看构建状态" (view build status), "在确认后触发构建" (trigger builds after confirmation) — which comprehensively cover the operate-an-existing-pipeline lifecycle this skill targets. It matches the anchor 'multiple specific concrete actions; comprehensive coverage' and is not score 4 because there is no meaningful coverage gap for the skill's stated scope.

5 / 5

Completeness

It explicitly answers both parts: 'what' — managing pipeline builds via MCP with four named operations — and 'when' — "当用户要操作现有流水线而不是修改代码实现时优先使用" (prefer when the user wants to operate existing pipelines rather than modify code). The when-clause is explicit and concrete with a clear boundary, matching the score-5 anchor; it is clearly above score 4 where the 'when' is only implicitly or loosely specified.

5 / 5

Trigger Term Quality

Natural terms a BK-CI user would say are well covered: "流水线" (pipeline), "构建历史" (build history), "构建状态" (build status), "触发构建" (trigger build), "启动参数" (startup parameters), plus the product name "BK-CI". A few natural synonyms are missing — e.g., "蓝盾" (the Chinese product name used in the body's title), "执行记录", or generic "CI/CD" phrasing — so it fits 'good keyword coverage; a few natural terms missing' rather than the comprehensive synonym coverage of score 5.

4 / 5

Distinctiveness Conflict Risk

The niche is distinct: MCP-driven operation of existing BK-CI pipeline builds, explicitly delimited against modifying code/pipeline definitions ("而不是修改代码实现"), which minimizes conflict with pipeline-design or code-editing skills. It matches 'clear niche with distinct triggers; minimal conflict risk' and does not fall to score 4, which requires noted overlap with closely related skills.

5 / 5

Total

19

/

20

Passed

Validation

100%

Checks the skill against the spec for correct structure and formatting. All validation checks must pass before discovery and implementation can be scored.

Validation — 16 / 16 Passed

Validation for skill structure

No warnings or errors.

Repository
TencentBlueKing/bk-ci
Reviewed

Table of Contents

Is this your skill?

If you maintain this skill, you can claim it as your own. Once claimed, you can manage eval scenarios, bundle related skills, attach documentation or rules, and ensure cross-agent compatibility.