CtrlK
BlogDocsLog inGet started
Tessl Logo

1k-monitor-pr-ci

Monitor OneKey PR checks and review threads, fix failures, address comments, and continue until the PR is ready.

60

Quality

71%

Does it follow best practices?

Run evals on this skill

Adds up to 20 points to the overall score

View guide

SecuritybySnyk

Low

Low-risk findings worth noting

Fix and improve this skill with Tessl

tessl review fix ./.skillshare/skills/1k-monitor-pr-ci/SKILL.md
SKILL.md
Quality
Evals
Security

Quality

Content

85%Weight 40%Scale 1-5

Reviews the quality of instructions and guidance provided to agents. Good implementation is clear, handles edge cases, and produces reliable results.

This is a highly actionable, well-sequenced orchestration skill with exemplary validation and error-recovery structure; nearly every instruction is an executable command. Its weaknesses are mild redundancy across sections and the total absence of progressive disclosure — all detail lives in one long file with no offloading to reference files.

Suggestions

Move the agent-check JSON field reference (schemaVersion fields list) and the REST/GraphQL fallback procedures into a references/ file, linked once from Step 1a, to cut main-file length.

Consolidate the GraphQL-only-resolve caveat and the auto-fix-without-asking rule so each is stated once (Step 1a/3a) rather than repeated in Important Notes and Error Handling.

Trim the Usage section to one or two representative invocations plus a note that both PR number and URL forms are accepted.

DimensionReasoningScore

Conciseness

The body is command-first and dense with executable specifics, but repeats a few rules in multiple places: the GraphQL-only thread-resolve limitation appears in Step 1a, Step 3d, and Error Handling; the auto-fix-without-asking rule appears in Step 3a and Important Notes; and Usage lists five near-identical invocations. Not 5 because this repetition could be consolidated; not 3 because there is no padding or explanation of concepts Claude already knows.

4 / 5

Actionability

Every step ships copy-paste-ready commands with exact flags: `yarn agent:check ... --json-file`, `gh run view <RUN_ID> --log-failed`, the REST reply endpoint, the full GraphQL resolveReviewThread mutation, and re-review requests with an API fallback. Fallback paths and edge cases (REST fallback, unsupported schemaVersion) are covered, matching the fully-executable/comprehensive anchor.

5 / 5

Workflow Clarity

Steps 0-4 are clearly sequenced with a state-to-action decision table (CI status x unresolved threads), per-iteration validation of PR state (abort if CLOSED/MERGED), pre-existing-failure verification against base, and explicit fix-to-push-to-re-poll feedback loops. The Error Handling section distinguishes blocking from non-blocking failures. This matches the anchor with explicit validation steps, feedback loops, and checklists.

5 / 5

Progressive Disclosure

There are no bundle files (references/, scripts/, assets/ do not exist) and no pointers to any; everything is inline in a single ~360-line SKILL.md. Section headers are good, so it is not a monolithic wall (not 2), but content that could live one level deep — the agent-check JSON field reference, the REST/GraphQL fallback procedures, and the display examples — is inlined, matching the 'some structure; content that should be separate is inline' anchor rather than 4.

3 / 5

Total

17

/

20

Passed

Description

58%Weight 40%Scale 1-5

Based on the skill's description, can an agent find and select it at the right time? Clear, specific descriptions lead to better discovery.

The description communicates concrete capabilities in third person with a clear purpose, but lacks an explicit 'Use when...' trigger clause and common trigger vocabulary like CI, GitHub, or pull request. It is distinct from sibling skills but would be more discoverable with trigger phrases.

Suggestions

Add an explicit trigger clause, e.g. "Use when the user asks to monitor a PR's CI checks or review comments, keep CI green, or drive a PR to merge-ready."

Include natural trigger keywords users would actually say: "CI", "GitHub pull request", "review comments", "fix CI failures".

Optionally name the reply/resolve-thread capabilities ("reply to reviewers and resolve threads") to make the what-coverage comprehensive.

DimensionReasoningScore

Specificity

"Monitor OneKey PR checks and review threads, fix failures, address comments, and continue until the PR is ready" lists several concrete actions (monitor, fix failures, address comments, iterate to done). Not 5 because it omits capabilities the body demonstrates (reply/resolve threads, polling interval); not 3 because it goes well beyond 1-2 actions.

4 / 5

Completeness

The "what" is clear (monitor checks and threads, fix failures, address comments, drive to ready), but there is no "Use when..." clause or equivalent explicit trigger guidance, which caps completeness at 3 per the judging guidelines. Not 2 because the "what" is concrete and specific.

3 / 5

Trigger Term Quality

"PR checks" and "review threads" are natural user phrases, but the description misses common variations users would say: "CI", "GitHub", "pull request", "review comments". Matches the anchor 'some relevant keywords but missing common variations or synonyms', not 4 since several natural terms are absent.

3 / 5

Distinctiveness Conflict Risk

"Monitor OneKey PR checks and review threads" carves a fairly distinct monitoring/fix-loop niche tied to the OneKey repo. Minor overlap risk with generic PR-review or CI-fix skills keeps it below 5; well above generic skill descriptions, so not 3.

4 / 5

Total

14

/

20

Passed

Validation

93%

Checks the skill against the spec for correct structure and formatting. All validation checks must pass before discovery and implementation can be scored.

Validation — 15 / 16 Passed

Validation for skill structure

CriteriaDescriptionResult

frontmatter_unknown_keys

Unknown frontmatter key(s) found; consider removing or moving to metadata

Warning

Total

15

/

16

Passed

Repository
OneKeyHQ/app-monorepo
Reviewed

Table of Contents

Is this your skill?

If you maintain this skill, you can claim it as your own. Once claimed, you can manage eval scenarios, bundle related skills, attach documentation or rules, and ensure cross-agent compatibility.