CtrlK
BlogDocsLog inGet started
Tessl Logo

veomni-develop

VeOmni-specific checklist for feature development and refactoring. Covers impact analysis across modalities, trainer hierarchy, data pipeline, and distributed code. Use before implementing any non-trivial change. For model-specific or ops-specific work, use veomni-new-model or veomni-new-op instead. Trigger: 'add feature', 'implement', 'refactor', 'reorganize', 'new capability'.

72

Quality

90%

Does it follow best practices?

Run evals on this skill

Adds up to 20 points to the overall score

View guide
SecuritybySnyk

Passed

No findings from the security scan

SKILL.md
Quality
Evals
Security

Quality

Content

92%Weight 40%Scale 1-5

Reviews the quality of instructions and guidance provided to agents. Good implementation is clear, handles edge cases, and produces reliable results.

An exemplary instruction-only skill body: dense with non-obvious project constraints, a validated refactoring workflow with baseline checkpoints, and clean delegation of detail to existing project knowledge files. The only minor gap is that a few directives ('grep all YAML files first') could be made copy-paste executable.

DimensionReasoningScore

Conciseness

Every line carries repo-specific knowledge Claude could not know (import-time MODELING_REGISTRY population, strict SP ordering 'pad → slice → FA kwargs → slice position_ids', position_ids == 0 segment boundaries, silent YAML config breakage). There is no padding and no explanation of concepts Claude already knows, matching anchor 5's 'every token earns its place'; anchor 4 would require trimmable over-explanation, which is absent.

5 / 5

Actionability

Guidance is mostly executable and concrete: exact paths (veomni/data/data_collator.py, .agents/knowledge/testing.md), commands ('run pytest tests/', 'grep all YAML files first'), and a precise collator ordering rule. It falls short of anchor 5 because some direction stays at the hint level (e.g. 'grep all YAML files first' without the actual grep command) and the impact-analysis table says what to check but not how.

4 / 5

Workflow Clarity

The Refactoring Safety Rules section is a clear sequence with explicit validation checkpoints and a feedback loop: 'Baseline first: run pytest tests/ before any change, record results', one change per commit with 'verify tests match baseline', and 'Check baseline again at the end — results must be identical'. This matches anchor 5's explicit validation with error-recovery guidance.

5 / 5

Progressive Disclosure

This is a compact (~57-line) single-file skill with no bundle directories; all sections are tight and well-organized, and detail is correctly delegated one level deep to existing project files (docs/, configs/, .agents/knowledge/*.md) with a compact inline summary of the testing rules. Per the rubric's short-skill exception, well-organized sections with no need for external references warrant a 5.

5 / 5

Total

19

/

20

Passed

Description

83%Weight 40%Scale 1-5

Based on the skill's description, can an agent find and select it at the right time? Clear, specific descriptions lead to better discovery.

A strong description that clearly states what the skill covers, when to use it, and how it differs from sibling skills, with concrete trigger phrases. The remaining gaps are minor: trigger synonyms are thin ('implement' alone is broad) and the capability description stays at the level of coverage areas rather than concrete operations.

Suggestions

Broaden trigger coverage with natural synonyms users would say, e.g. add 'change', 'modify', 'extend', or 'touching the trainer' alongside the current five triggers.

Tighten the 'implement' trigger (e.g. 'implement a feature', 'non-trivial implementation') so the description is less likely to fire on small routine coding tasks outside the checklist's intent.

Make the core action slightly more concrete: instead of 'checklist for feature development', name the primary action performed (e.g. 'run an impact analysis before implementing').

DimensionReasoningScore

Specificity

The description names the domain and lists several concrete coverage areas ("impact analysis across modalities, trainer hierarchy, data pipeline, and distributed code"), matching anchor 4's 'several specific actions; minor gaps'. It falls short of anchor 5 because the actions remain one level abstract ('checklist', 'impact analysis') rather than fully concrete operations.

4 / 5

Completeness

Both questions are explicitly answered: what ("VeOmni-specific checklist for feature development and refactoring. Covers impact analysis across modalities, trainer hierarchy, data pipeline, and distributed code") and when ("Use before implementing any non-trivial change" plus concrete trigger phrases). This matches anchor 5 exactly; anchor 4 would require a less explicit 'when'.

5 / 5

Trigger Term Quality

Explicit triggers ('add feature', 'implement', 'refactor', 'reorganize', 'new capability') are natural phrases a user would say, giving good keyword coverage. Not anchor 5 because common synonyms and variations ('change', 'modify', 'extend') are missing, leaving a few natural terms uncovered.

4 / 5

Distinctiveness Conflict Risk

The "VeOmni-specific" scoping plus explicit routing to sibling skills ("For model-specific or ops-specific work, use veomni-new-model or veomni-new-op instead") gives a clear niche with minimal internal conflict. It stops short of anchor 5 because the bare trigger 'implement' is broad enough to fire on general coding tasks, creating minor overlap risk with non-VeOmni skills.

4 / 5

Total

17

/

20

Passed

Validation

100%

Checks the skill against the spec for correct structure and formatting. All validation checks must pass before discovery and implementation can be scored.

Validation — 16 / 16 Passed

Validation for skill structure

No warnings or errors.

Repository
ByteDance-Seed/VeOmni
Reviewed

Table of Contents

Is this your skill?

If you maintain this skill, you can claim it as your own. Once claimed, you can manage eval scenarios, bundle related skills, attach documentation or rules, and ensure cross-agent compatibility.