CtrlK
BlogDocsLog inGet started
Tessl Logo

embedded-captions

Add captions or subtitles to an existing single-subject talking-head video without editing the footage. Use for plain verbatim captions, cinematic captions embedded behind the subject, VFX captions, “炸/特效/酷炫字幕,” or a named identity from the 35-style catalog. Route by visual identity, not by backend engine. The quiet `anchor` rail is the default; embed every word only when the user explicitly wants a fully cinematic treatment. The workflow runs locally end to end, including transcription and subject matting; split multi-shot footage before applying it.

70

Quality

88%

Does it follow best practices?

Run evals on this skill

Adds up to 20 points to the overall score

View guide
SecuritybySnyk

Passed

No findings from the security scan

SKILL.md
Quality
Evals
Security

Quality

Content

77%Weight 40%Scale 1-5

Reviews the quality of instructions and guidance provided to agents. Good implementation is clear, handles edge cases, and produces reliable results.

The body is an unusually actionable and validation-heavy instruction set — exact commands, numeric thresholds, compile/render gates, and a preview-before-render QA loop with checklists — with excellent reference-table navigation in principle. Its two real weaknesses are redundancy (the pipeline, decision-gate probes, and core warnings are each stated multiple times, and the DNA registry is inlined despite having its own file) and bundle integrity: the most-cited files (CATALOG.md, themes/README.md, dna/README.md) are absent from the bundle, leaving the skill's central routing and schema steps unresolvable.

Suggestions

Merge 'Operational flow (TL;DR)' and 'Pipeline — 5 steps' into one canonical numbered sequence, state each load-bearing rule once (the 'embedding every word is the common mistake' warning appears four times), and drop either the inline 10-row DNA table or the dna/README.md pointer so the registry lives in exactly one place.

Ship CATALOG.md, themes/README.md, and dna/README.md in the bundle (or inline their essential content): they are referenced as the 'single source of truth for routing', the exact theme.json schema, and the DNA decision rule, yet do not exist — the skill's step 1 and both mode-authoring steps currently dead-end.

Deduplicate the pre-flight material between 'Decision gate — RUN FIRST' and 'Pre-flight probes' into a single probe list, and move the per-identity register/scene-fit details of the DNA table into dna/README.md, keeping only the pick rule inline.

DimensionReasoningScore

Conciseness

The body is dense with unique domain knowledge (no filler, no explanations of concepts Claude already knows), but it is noticeably duplicated rather than lean: the 5-step pipeline is stated twice ('Operational flow (TL;DR)' and 'Pipeline — 5 steps'), the decision-gate probes appear in two sections, 'embedding every word is the common mistake' is repeated four times, and a 10-row DNA registry table is inlined alongside a dedicated dna/README.md. This fits anchor 3 ('mostly efficient but could be tightened') better than 4, where only minor trimming would be needed — merging the duplicated sections would cut substantial length. It is above anchor 2 because the verbosity is structural repetition of load-bearing rules, not padded or generic explanation.

3 / 5

Actionability

Fully executable throughout: copy-paste commands ('bash scripts/prepare.sh <project>', 'node scripts/preview-frames.cjs <project>', exact ffprobe/ffmpeg probe commands with flags, pinned 'npm install ... sharp@0.35.3 puppeteer@25.8.0 gsap@3.15.0'), exact numeric thresholds (luma > 150, ≥30% face uncovered per 0.3s, 80ms timing tolerance, 0.5s minimum on-screen), and pointers to exact schemas ('Schema: scripts/make-cinematic.cjs header'). Matches the anchor-5 pattern of copy-paste-ready commands covering the common cases; anchor 4's 'minor gaps' does not apply.

5 / 5

Workflow Clarity

The sequence is explicit and numbered (init → prepare → author JSON per mode → preview QA → render, with a 'Decision gate — RUN FIRST' before either mode), validation is explicit at every stage (compile-time verbatim gate, check-timing.cjs --strict, render gates for timing/occlusion/overflow/hand-off, a 5-point Visual QA failure checklist plus 5 positive checks), and there is a genuine feedback loop ('Apply fixes in plan.json / theme.json, recompile, re-preview ... Render once, when the previews pass', plus the fresh-eyes subagent review). This matches anchor 5's 'explicit validation steps; feedback loops for error recovery; checklists'; anchor 4 is ruled out because no checkpoint is missing.

5 / 5

Progressive Disclosure

Structure is genuinely good — a Shared knowledge table mapping 14 reference docs, per-need pointers ('skim by need', 'read before embedding') — but scored against the actual bundle: the three most load-bearing targets are missing. CATALOG.md (referenced ~10 times, called 'single source of truth for routing'), themes/README.md ('read FIRST — ... the exact theme.json schema'), and dna/README.md ('has the decision rule') do not exist in the bundle, so the routing and theme-schema steps dead-end; the inline 10-row DNA table also duplicates the (missing) registry file. This is worse than anchor 4's 'minor organization gaps' but better than anchors 1–2 (no deep nesting, no monolith; references are clearly signaled), so anchor 3 fits.

3 / 5

Total

16

/

20

Passed

Description

95%Weight 40%Scale 1-5

Based on the skill's description, can an agent find and select it at the right time? Clear, specific descriptions lead to better discovery.

A strong description: third-person, concrete about what it does (captions/subtitles on existing footage, transcription and matting run locally), with an explicit and specific 'Use for' clause covering natural trigger terms in both English and Chinese, plus useful boundary guidance (single-subject, split multi-shot footage, anchor rail default). The only mild weakness is that the action enumeration is one core capability plus pipeline detail rather than several distinct actions.

DimensionReasoningScore

Specificity

Concrete actions are named — 'Add captions or subtitles to an existing single-subject talking-head video without editing the footage' plus 'runs locally end to end, including transcription and subject matting' — covering the domain with only minor gaps (compositing/preview/render steps are only implied by 'workflow'). It sits at anchor 4 rather than 5 because the action list is one core action plus supporting pipeline capabilities, slightly less comprehensive than the anchor-5 pattern of several distinct enumerated actions.

4 / 5

Completeness

Both questions answered explicitly: what — 'Add captions or subtitles to an existing single-subject talking-head video without editing the footage'; when — 'Use for plain verbatim captions, cinematic captions embedded behind the subject, VFX captions, 炸/特效/酷炫字幕, or a named identity from the 35-style catalog.' This matches the anchor-5 good example structure ('what' + explicit 'Use when/for' with concrete trigger phrases), and is above anchor 4 where the 'when' is merely present but less specific.

5 / 5

Trigger Term Quality

Natural terms are comprehensively covered with synonyms across languages: 'captions', 'subtitles', 'plain verbatim captions', 'cinematic captions embedded behind the subject', 'VFX captions', and the Chinese variants '炸/特效/酷炫字幕' that a user would actually say. Not below 5: no common synonym family for this task is missing; not applicable above 5.

5 / 5

Distinctiveness Conflict Risk

A clear niche — captioning existing single-subject talking-head footage with the footage untouched — with distinct triggers (verbatim/cinematic/VFX/Chinese VFX terms) and an explicit routing disambiguator ('Route by visual identity, not by backend engine'). Minimal overlap risk with generic video-editing or composition skills; anchor 5 fits, and anchor 4 (minor overlap with closely related skills) is weaker since the 'without editing the footage' boundary separates it from editing skills.

5 / 5

Total

19

/

20

Passed

Validation

93%

Checks the skill against the spec for correct structure and formatting. All validation checks must pass before discovery and implementation can be scored.

Validation — 15 / 16 Passed

Validation for skill structure

CriteriaDescriptionResult

relative_links

Relative link issues: 12 missing, 1 suspicious

Warning

Total

15

/

16

Passed

Repository
heygen-com/hyperframes
Reviewed

Table of Contents

Is this your skill?

If you maintain this skill, you can claim it as your own. Once claimed, you can manage eval scenarios, bundle related skills, attach documentation or rules, and ensure cross-agent compatibility.