CtrlK
BlogDocsLog inGet started
Tessl Logo

hyperframes-cli

Use the HyperFrames CLI development loop: init, add, catalog, capture, lint, check, snapshot, compare, grade-compare, preview, play, present, beats, keyframes, single or batch render, publish, cloud, cloudrun, feedback, lambda, doctor, browser, info, upgrade, skills, compositions, timeline, history, clean, docs, benchmark, telemetry, transcribe, auth, tts, and remove-background. Also use when diagnosing build or render failures. validate, inspect, and layout are deprecated aliases; use check. Covers local, HeyGen-hosted cloud, AWS Lambda, and Google Cloud Run rendering.

69

Quality

87%

Does it follow best practices?

Run evals on this skill

Adds up to 20 points to the overall score

View guide
SecuritybySnyk

Low

Low-risk findings worth noting

SKILL.md
Quality
Evals
Security

Quality

Content

88%Weight 40%Scale 1-5

Reviews the quality of instructions and guidance provided to agents. Good implementation is clear, handles edge cases, and produces reliable results.

An exceptionally actionable skill body: every workflow step is tied to an executable command, validation gates, or a verified one-level-deep reference file, and the development loop has explicit checkpoints and feedback loops. The only weaknesses are verbosity in a few deep-dive passages and reference-grade detail inlined in the overview that would sit better in the reference files.

Suggestions

Move the catalog tier mechanics (the dropped/unindexed skew semantics, refetch remedies, and report_gap workflow) into a dedicated reference file and keep a one-line decision rule inline, tightening the longest bullet in Agent conventions.

Move the usage-window result schema (known/unknown fields, per-harness window descriptions) into references/upgrade-info-misc.md, keeping only the command invocation and reporting rule in the body.

The feedback reproduction-packet format is described at length in the body while also living in references/preview-render.md — keep only the consent/privacy warning and pointer inline.

DimensionReasoningScore

Conciseness

The body is dense and command-first with essentially no explanation of concepts Claude already knows — everything is CLI-specific contract. However, a few passages carry reference-grade detail inline that could be trimmed or moved out, e.g. the tier-skew mechanics paragraph ("dropped counts ranked names this registry cannot install... Both counts are of names rather than of results, so either can exceed total") and the usage-window schema enumeration ("status: "known"", "harness", "planTier", "session", and "weekly"...). This matches the score-4 anchor ("efficient; minor instances of over-explanation that could be trimmed"), not score 5, since some sections run long even though accurate.

4 / 5

Actionability

Nearly every section carries copy-paste-ready executable commands: the verification block ("npx hyperframes check\nnpx hyperframes preview --background\nnpx hyperframes render --quality looks --output out.mp4\ntest -s out.mp4\nffprobe -v error -show_format -show_streams out.mp4"), the render-choices table mapping each need to a full command, the doctor gate ("npx hyperframes doctor --json | jq -e '.ok' >/dev/null"), and the feedback command with parameters. Placeholders like <t1>,<t2>,<t3> and <your-name> are legitimate runtime parameters, matching the score-5 anchor for fully executable coverage of common cases.

5 / 5

Workflow Clarity

The 10-step development loop is explicitly sequenced with validation checkpoints throughout: lint after the first HTML pass (step 4), the final "check" gate (step 5), HTTP 200 verification of the preview URL (step 7), render only after approval (step 8), and output verification via test -s and ffprobe duration comparison (step 10). Feedback loops for recovery are present (history undo exits 2 on conflict and prints both choices; sub-composition smoke test treats mount defects as render-blocking). Batch render and the destructive undo both have documented validation, so the batch/destructive cap at 3 does not apply. This matches the score-5 anchor (clear sequence, explicit validation, feedback loops).

5 / 5

Progressive Disclosure

Structure is strong: a mandatory command-to-reference table maps every command family to one of ten references (all verified to exist and be substantive: beats.md, cloud.md, cloudrun.md, compare-and-batch.md, doctor-browser.md, init-and-scaffold.md, lambda.md, lint-validate-inspect.md, preview-render.md, upgrade-info-misc.md), one level deep with clear headers and cross-skill pointers. But the overview body itself inlines a fair amount of reference-grade detail (catalog tier semantics and dropped/unindexed skew mechanics, the usage/harness result schema, the feedback reproduction-packet format), which keeps it below the score-5 anchor's "clear overview with content appropriately split". It is above the score-3 anchor because references are clearly signaled and the split is otherwise appropriate.

4 / 5

Total

18

/

20

Passed

Description

83%Weight 40%Scale 1-5

Based on the skill's description, can an agent find and select it at the right time? Clear, specific descriptions lead to better discovery.

A strong description that comprehensively enumerates the CLI's concrete commands and rendering targets with an explicit diagnostics trigger and clear deprecation guidance. The primary authoring use-case trigger and natural domain synonyms (video, motion, composition) are the only gaps keeping it from top marks.

DimensionReasoningScore

Specificity

The description enumerates the full concrete command surface — "init, add, catalog, capture, lint, check, snapshot, compare, grade-compare, preview, play, present, beats, keyframes, single or batch render, publish, cloud, cloudrun, feedback, lambda, doctor..." — plus deployment targets ("local, HeyGen-hosted cloud, AWS Lambda, and Google Cloud Run"). This is comprehensive coverage of specific actions, matching the score-5 anchor rather than the score-4 anchor, which expects only "several specific actions" with gaps.

5 / 5

Completeness

The "what" is explicit and comprehensive (the full development loop with every command), and a "when" clause exists ("Also use when diagnosing build or render failures"). But the primary authoring trigger — when to use the skill for creating/editing a video project — is only implied, so the "when" could be more explicit, matching the score-4 anchor. It is not score 5 because the main use case lacks explicit concrete trigger phrases, and the score-3 cap (missing use-when) does not apply since one is stated.

4 / 5

Trigger Term Quality

Natural trigger terms are present — "diagnosing build or render failures", "render", "publish", "preview", "transcribe", "tts", "remove-background" — phrasing users of this CLI would actually say. However, common domain synonyms are missing ("video", "animation", "motion", "composition authoring") and no file extensions are given, so it falls short of the score-5 anchor's "comprehensive coverage including synonyms and file extensions" while exceeding the score-3 anchor's "some relevant keywords but missing common variations".

4 / 5

Distinctiveness Conflict Risk

The description is anchored to a named niche tool ("HyperFrames CLI") with distinct command triggers and even disambiguates deprecated aliases ("validate, inspect, and layout are deprecated aliases; use check"), minimizing both cross-skill and within-tool confusion. This matches the score-5 anchor's "clear niche with distinct triggers; minimal conflict risk".

5 / 5

Total

18

/

20

Passed

Validation

87%

Checks the skill against the spec for correct structure and formatting. All validation checks must pass before discovery and implementation can be scored.

Validation — 14 / 16 Passed

Validation for skill structure

CriteriaDescriptionResult

relative_links

Relative link issues: 1 suspicious

Warning

referenced_paths_exist

Referenced path issues: 2 missing

Warning

Total

14

/

16

Passed

Repository
heygen-com/hyperframes
Reviewed

Table of Contents

Is this your skill?

If you maintain this skill, you can claim it as your own. Once claimed, you can manage eval scenarios, bundle related skills, attach documentation or rules, and ensure cross-agent compatibility.