CtrlK
BlogDocsLog inGet started
Tessl Logo

imagegen

Generate or edit images via BlockRun's image API. Trigger when the user asks to generate, create, draw, make an image — or to edit, modify, change, or retouch an existing image.

65

Quality

82%

Does it follow best practices?

Run evals on this skill

Adds up to 20 points to the overall score

View guide
SecuritybySnyk

Passed

No findings from the security scan

SKILL.md
Quality
Evals
Security

Image Generation & Editing

Generate or edit images through ClawRouter. Payment is automatic via x402.

Shortcuts:

  • Slash: /cr-imagegen <prompt> [--model=<alias>] [--size=1024x1024] [--n=1] (/imagegen still accepted in chat for backward compatibility)
  • Partner tool: blockrun_image_generation (LLM-callable) / blockrun_image_edit (inpainting)

Generate an Image

POST to http://localhost:8402/v1/images/generations:

{
  "model": "google/nano-banana",
  "prompt": "a golden retriever surfing on a wave",
  "size": "1024x1024",
  "n": 1
}

Response:

{
  "created": 1741460000,
  "data": [{ "url": "http://localhost:8402/images/abc123.png" }]
}

Display inline: ![generated image](http://localhost:8402/images/abc123.png)

Model Selection

AliasFull IDPriceSizesBest for
nano-bananagoogle/nano-banana$0.051024×1024Default — fast, cheap, good quality
banana-2google/nano-banana-2$0.091024×1024Gemini 3.1 Flash imagegen — sharper than nano-banana at 1K
banana-progoogle/nano-banana-pro$0.10–$0.151024×1024, 2048×2048, 4096×4096High-res, large format
gpt-imageopenai/gpt-image-1$0.02–$0.041024×1024, 1536×1024, 1024×1536Budget option; supports editing
gpt-image-2openai/gpt-image-2$0.06–$0.121024×1024, 1536×1024, 1024×1536Photorealistic, reasoning-driven, text rendering (slow — proxy polls up to 5min); legacy dalle alias routes here
flareopenai/gpt-image-2.5-flare$0.28–$0.561024×1024, 1536×1024, 1024×1536GPT Image 2.5, fast — top OpenAI quality, caller can pick quality at a flat price
sunburstopenai/gpt-image-2.5-sunburst$0.28–$0.561024×1024, 1536×1024, 1024×1536GPT Image 2.5, precision — best for high-fidelity edits
seedreambytedance/seedream-5-pro$0.045–$0.09up to 2848×1600 / 2304×1728Flagship quality, reference-image support
grok-imaginexai/grok-imagine-image$0.021024×1024xAI Grok image style
grok-imagine-2xai/grok-imagine-image-2.0$0.041024×1024Grok Imagine 2.0 — between grok-imagine and pro
grok-imagine-proxai/grok-imagine-image-pro$0.071024×1024Grok high-quality
cogviewzai/cogview-4$0.015–$0.02512×512 to 1440×1440Cheapest — Zhipu CogView

Choosing a model:

  • Default → nano-banana
  • "high res" / "large" → banana-pro
  • "photorealistic" / complex scenes → gpt-image-2
  • "flagship quality" / reference image → seedream
  • "budget" / "cheap" → cogview
  • "top quality" / "best OpenAI" → flare (fast) or sunburst (precision)
  • "editable" / "inpainting" → gpt-image (cheapest), gpt-image-2, or sunburst (most precise)
  • "artistic" / flexible content → grok-imagine
  • "grok style" → grok-imagine or grok-imagine-pro

Choosing a size:

  • Default: 1024x1024 (the only size every model accepts)
  • Portrait: 1024x1536 (gpt-image / gpt-image-2 / flare / sunburst) or 1728x2304 (seedream)
  • Landscape: 1536x1024 (gpt-image / gpt-image-2 / flare / sunburst), 1344x768 (cogview), or 2048x1024 / 1280x720 (seedream)
  • High-res: 2048x2048 / 4096x4096 with banana-pro; 2848x1600 with seedream
  • The gateway validates size per model BEFORE payment and rejects unknown ones — do not invent sizes outside each model's list above

Edit an Existing Image

POST to http://localhost:8402/v1/images/image2image:

{
  "model": "openai/gpt-image-1",
  "prompt": "make the background a snowy mountain landscape",
  "image": "https://example.com/photo.jpg",
  "size": "1024x1024",
  "n": 1
}

ClawRouter automatically downloads URLs and reads local file paths — pass them directly, no manual base64 conversion needed.

Optional mask field: a second image (URL or path) that marks which areas to edit (white = edit, black = keep).

Response is identical to generation:

{
  "created": 1741460000,
  "data": [{ "url": "http://localhost:8402/images/xyz456.png", "revised_prompt": "..." }]
}

Supported models for editing: openai/gpt-image-1 (default, $0.02), openai/gpt-image-2 ($0.06), openai/gpt-image-2.5-sunburst ($0.28), google/nano-banana ($0.05), google/nano-banana-2 ($0.09), google/nano-banana-pro ($0.10). mask works with the OpenAI models only. openai/gpt-image-2.5-flare cannot edit.


Example Interactions

User: Draw me a cyberpunk city at night → POST to /v1/images/generations, model nano-banana, prompt as given.

User: Generate a high-res portrait of a samurai → POST to /v1/images/generations, model seedream, size 1728x2304.

User: Edit this photo to add a sunset background: https://example.com/portrait.jpg → POST to /v1/images/image2image, model gpt-image, image = the URL, prompt = "add a warm sunset background".

User: Change the background in my image to a beach (attaches local file) → POST to /v1/images/image2image, image = the local file path, prompt describes the change.


Notes

  • Payment is automatic, and which rail it uses depends on how ClawRouter was started: an x402 USDC micropayment from the user's wallet, or a draw on account credit if a BlockRun API key is configured. Either way the agent does nothing.
  • If the call fails with a payment error, check GET http://localhost:8402/health and read authMode before telling the user how to fix it: wallet → fund the wallet at blockrun.ai; api-key → top up account credit at user.blockrun.ai/dashboard/credits. Naming the wrong one sends the user to a page that cannot fix their error.
  • Google models may return base64 internally — ClawRouter uploads automatically and returns a hosted URL
  • OpenAI image models enforce OpenAI content policy; use nano-banana or grok-imagine for more flexibility
  • Image editing works with gpt-image-1/2, sunburst and the nano-banana models (not flare, seedream, grok or cogview); generation supports all listed models
Repository
BlockRunAI/ClawRouter
Last updated
First committed

Is this your skill?

If you maintain this skill, you can claim it as your own. Once claimed, you can manage eval scenarios, bundle related skills, attach documentation or rules, and ensure cross-agent compatibility.