CtrlK
BlogDocsLog inGet started
Tessl Logo

infinitetalk

自媒体创作者与内容创作者在制作数字人播报或视频配音时,只需输入单张人像与音频,即可自动生成唇形、表情、动作完美同步的无限时长说话视频。一键打造高质量虚拟主播内容,告别繁琐拍摄,让音视频创作更高效!

62

Quality

73%

Does it follow best practices?

Run evals on this skill

Adds up to 20 points to the overall score

View guide

SecuritybySnyk

Passed

No findings from the security scan

Fix and improve this skill with Tessl

tessl review fix ./skills/infinitetalk/SKILL.md
SKILL.md
Quality
Evals
Security

Quality

Content

86%Weight 40%Scale 1-5

Reviews the quality of instructions and guidance provided to agents. Good implementation is clear, handles edge cases, and produces reliable results.

The content is a well-structured, actionable overview: concrete script invocations with copy-paste examples, clear sequenced workflows with validation/feedback, and clean one-level-deep references to real bundle files. The only notable gap is minor validation lightness in the secondary modes.

DimensionReasoningScore

Conciseness

The body is efficient and assumes Claude's competence — it does not explain what a video or a model is — but the '任务目标' capability list partly restates the operating steps, leaving minor instances that could be trimmed; it is above 'mostly efficient with some unnecessary explanation' (3) but short of 'every token earns its place' (5).

4 / 5

Actionability

It provides the concrete script path, documented parameters (--input_path, --size, --mode, --sample_audio_guide_scale), and three copy-paste-ready bash examples covering the common cases (basic image-to-video, long video via streaming, TTS), matching the 'fully executable, copy-paste ready, covers common cases' anchor.

5 / 5

Workflow Clarity

Each mode is clearly sequenced (prepare input → run generation → verify output), and Mode 1 includes an explicit validation step plus a feedback loop ('如有异常,调整 sample_audio_guide_scale') with additional OOM recovery guidance in 注意事项; it is not a 5 because Modes 2 and 3 have lighter validation checkpoints.

4 / 5

Progressive Disclosure

The body is an overview with a dedicated 资源索引 section and well-signaled one-level-deep references to real bundle files (references/model_download.md, environment_setup.md, usage_examples.md, scripts/infer_infinitetalk.py), with detail appropriately split out and easy to navigate.

5 / 5

Total

18

/

20

Passed

Description

61%Weight 40%Scale 1-5

Based on the skill's description, can an agent find and select it at the right time? Clear, specific descriptions lead to better discovery.

The description conveys a clear, distinctive capability with good natural trigger terms, but it is written as marketing prose with over-claims and lacks an explicit 'Use when' trigger clause, which caps completeness and weakens specificity.

Suggestions

Add an explicit 'Use when...' trigger clause (e.g., '当需要生成音频驱动的数字人播报、视频配音或虚拟主播内容时使用') to lift completeness above 3.

Rewrite in third-person voice ('从单张人像与音频生成唇形/表情/动作同步的说话视频') and drop marketing fluff/over-claims such as '完美同步', '一键打造', and '告别繁琐拍摄' to tighten specificity.

DimensionReasoningScore

Specificity

Names the domain and concrete actions (input a single portrait + audio to auto-generate an infinite-duration talking video with lip/expression/motion sync), but pads them with marketing fluff and over-claims ('完美同步', '一键打造', '告别繁琐拍摄') and uses scenario/second-person voice rather than clean third person, so it does not reach the 'several specific actions, minor gaps' anchor.

3 / 5

Completeness

It gives a clear 'what' (generate a synchronized talking video) and a scenario-style 'when' ('自媒体创作者与内容创作者在制作数字人播报或视频配音时'), but there is no explicit 'Use when...' trigger clause, which per the guidelines caps completeness at 3.

3 / 5

Trigger Term Quality

Good coverage of natural terms users would actually say — '数字人播报', '视频配音', '虚拟主播', '说话视频' — with only minor synonyms missing; it sits above 'some relevant keywords' (3) but lacks the exhaustive synonym/extension coverage of a 5.

4 / 5

Distinctiveness Conflict Risk

The audio-driven talking-video / digital-human niche is specific with distinct triggers and low conflict risk; it is not a 5 only because there is minor overlap risk with general image-to-video or TTS skills.

4 / 5

Total

14

/

20

Passed

Validation

93%

Checks the skill against the spec for correct structure and formatting. All validation checks must pass before discovery and implementation can be scored.

Validation15 / 16 Passed

Validation for skill structure

CriteriaDescriptionResult

frontmatter_unknown_keys

Unknown frontmatter key(s) found; consider removing or moving to metadata

Warning

Total

15

/

16

Passed

Repository
anbeime/skill
Reviewed

Table of Contents

Is this your skill?

If you maintain this skill, you can claim it as your own. Once claimed, you can manage eval scenarios, bundle related skills, attach documentation or rules, and ensure cross-agent compatibility.