CtrlK
BlogDocsLog inGet started
Tessl Logo

openai-image-gen

OpenAI Images API: batches, prompt sampler, gallery.

54

Quality

60%

Does it follow best practices?

Run evals on this skill

Adds up to 20 points to the overall score

View guide

SecuritybySnyk

Passed

No findings from the security scan

Fix and improve this skill with Tessl

tessl review fix ./skills/openai-image-gen/SKILL.md
SKILL.md
Quality
Evals
Security

Quality

Content

80%Weight 40%Scale 1-5

Reviews the quality of instructions and guidance provided to agents. Good implementation is clear, handles edge cases, and produces reliable results.

The body is an exemplar of token efficiency: terse, correct, and immediately actionable, with flags that match the real script. Its main weaknesses are hardcoded machine-specific paths (hurting portability/copy-paste readiness) and the absence of any validation checkpoint before a batch of paid API calls, despite the script offering --dry-run for exactly that purpose.

Suggestions

Replace absolute personal paths with a skill-relative invocation (e.g., "python3 scripts/gen.py relative to this skill's directory") so commands are copy-paste ready on any machine.

Add a validation checkpoint for the batch flow: "Preview prompts first with --dry-run; then run for real" — this both caps cost and gives a feedback loop before N API calls.

Document the script's remaining useful options and failure behavior (--timeout, --sleep, OPENAI_BASE_URL override, and what the error output looks like) so Claude can diagnose failures without reading the source.

DimensionReasoningScore

Conciseness

The body is lean and efficient — a one-line purpose statement, an env-var requirement, runnable commands, example flag invocations, and a three-item output listing, with zero padding or explanation of concepts Claude already knows. Every token earns its place, matching anchor 5; there is nothing extraneous to trim toward anchor 4.

5 / 5

Actionability

Commands are concrete and executable, and the flags shown (--count, --model, --size, --quality, --out-dir, --prompt) all match the actual script interface. It falls short of anchor 5 because paths are hardcoded to the author's environment ("python3 ~/Projects/agent-scripts/skills/openai-image-gen/scripts/gen.py") rather than being skill-relative or copy-paste ready elsewhere, and useful script capabilities like --dry-run and error/timeout behavior are omitted.

4 / 5

Workflow Clarity

The sequence (run script, open the generated gallery) is clear and the script path is concrete, but this is a batch operation — it fires N paid API calls — and the body includes no validation or verification checkpoint (e.g., running --dry-run first to preview prompts, or confirming outputs before opening the gallery). Per the judging guidelines, a batch skill without validation is capped at 3 even if single-purpose; it scores above anchor 2 because steps are concrete and unambiguous rather than vague.

3 / 5

Progressive Disclosure

This is a short single-purpose skill (~36 lines) whose only bundle file is scripts/gen.py, which exists and is what the Run section invokes; no external references are needed. The Setup / Run / Output sections are well-organized and appropriately sized for the SKILL.md overview role, matching the simple-skill exception for a top score.

5 / 5

Total

17

/

20

Passed

Description

41%Weight 40%Scale 1-5

Based on the skill's description, can an agent find and select it at the right time? Clear, specific descriptions lead to better discovery.

The description names a specific, distinctive domain but is a fragment list rather than a statement of capabilities: no action verbs, no natural trigger phrases like "generate images", and no "Use when..." guidance. It is decipherable but under-serves both discovery (triggering) and comprehension of what the skill actually does.

Suggestions

Add explicit trigger guidance, e.g. "Use when the user wants to generate images via the OpenAI Images API, create image batches, or build a prompt-driven image gallery."

Lead with third-person action verbs stating what the skill does, e.g. "Generates batches of images from sampled or user-supplied prompts via the OpenAI Images API and builds a thumbnail gallery of the results."

Include natural trigger terms and synonyms users would say ("generate images", "image generation", "DALL-E", "gpt-image") rather than only the API name.

DimensionReasoningScore

Specificity

The description "OpenAI Images API: batches, prompt sampler, gallery" names the domain but lists feature nouns ("batches, prompt sampler, gallery") rather than concrete actions — there is no verb phrase at all, matching anchor 2 ('names the domain but actions are minimal or generic'). It is below anchor 3, which requires 1-2 concrete actions (e.g., 'generates images'), and above anchor 1 because the domain is named specifically.

2 / 5

Completeness

A partial 'what' can be deciphered (image generation in batches producing a gallery), but the 'when' is completely absent — there is no "Use when..." clause or equivalent trigger guidance, which caps completeness at 3 per the judging guidelines. It is not 4 because 'when' guidance is missing entirely, not merely imprecise; not 2 because the 'what', while terse, is decipherable and domain-specific.

3 / 5

Trigger Term Quality

"OpenAI Images API" is a relevant keyword, but the natural phrases users would actually say — "generate images", "image generation", "create an image", "DALL-E" — are entirely absent; "batches", "prompt sampler", and "gallery" are weak trigger terms. This matches anchor 2 ('one or two relevant keywords; missing the natural phrases users say') rather than anchor 3, which requires broader synonym coverage.

2 / 5

Distinctiveness Conflict Risk

Naming the specific API ("OpenAI Images API") carves out a clear niche, so conflict risk with unrelated skills is low — matching anchor 4 ('mostly distinct; minor overlap risk with closely related skills'). It is below anchor 5 because the weak trigger phrasing leaves selection ambiguous against closely related image-generation skills.

4 / 5

Total

11

/

20

Passed

Validation

100%

Checks the skill against the spec for correct structure and formatting. All validation checks must pass before discovery and implementation can be scored.

Validation — 16 / 16 Passed

Validation for skill structure

No warnings or errors.

Repository
steipete/agent-scripts
Reviewed

Table of Contents

Is this your skill?

If you maintain this skill, you can claim it as your own. Once claimed, you can manage eval scenarios, bundle related skills, attach documentation or rules, and ensure cross-agent compatibility.