Node reference

Image Generator

Add Node → Generate → Image Generator

This is the node you’ll use the most. It makes still images — either from a text prompt on its own, or from a prompt plus reference images.

It arrives already wired to an empty image slot, ready to fill once you generate.

What’s on the node

The Image Generator node with its prompt box, model picker and settings

Area What goes in it
Prompt box Where you describe the image. Type into it directly, or wire something into the prompt handle and it turns read-only, showing the upstream text live.
Model and settings The model picker, and under it whatever that model exposes — resolution, aspect ratio, quality. The controls change with the model, and switching resets them.
P One only. Takes a Prompt Generator, a Text Area, a Prompt Split, or a finished result’s own prompt handle.
I Numbered 1, 2, 3… for reference images, in the order they’re sent to the model. Each takes an image, a Media Group or a Media Placeholder. Fill one and a fresh empty I appears below for the next.

A ring is hollow until something is connected and fills with colour once it is, so you can see at a glance what the node still needs. There’s more on all of this in How nodes connect.

Text-only, or with references?

Type a prompt and generate, and you get a plain text-to-image result.

Connect one or more images instead, and the node switches to image-to-image: now the model looks at your references while it generates, instead of working from words alone. This is the best way to keep something consistent across generations — a character, a product, a location — rather than hoping a text description gets it right.

Once you’ve connected references, you can call them out directly in the prompt, something like “combine the product from image 2 into image 1.”

Which model should you use?

Type Model Resolution Input Price range What it’s good at
Pro image-to-image
text-to-image
Nano Banana Pro 1K · 2K · 4K 10 $0.14–0.24 Google’s most capable, and the best here at plain-language prompting with references. Reach for it when consistency matters.
Pro image-to-image
text-to-image
GPT Image 2 1K · 2K · 4K 10 $0.02–0.73 The node’s default. A solid all-rounder for both text-to-image and image-to-image.
Pro image-to-image
text-to-image
Nano Banana 2 1K · 2K · 4K 10 $0.045–0.14 Same family, a step down. Still excellent with references and simple prompts.
Good image-to-image
text-to-image
Nano Banana HD 10 $0.038 The original Google model. Solid for straightforward edits, and cheap enough to run test after test at good quality.
Good image-to-image
text-to-image
Seedream 5.0 Pro 1K · 2K 10 $0.045–0.09 Strong image-to-image, for more control over composition and style.
Creative image-to-image
text-to-image
Flux 2.0 HD 3 $0.07 Fewer reference slots, more character in the result. For a look, not a faithful copy.
Creative image-to-image
text-to-image
Z-Image HD 1 $0.005 One reference, and the cheapest run here. Good for quick variations.
Creative text-to-image Ideogram 4.0 1K · 2K 0 $0.025–0.20 Text-only, and the best choice when you need clean, accurate text inside the image.
Creative text-to-image ERNIE HD 0 $0.03 Text-only. Worth trying for a different feel from a prompt alone.

Prices are indicative and set by the providers, who revise them: see wavespeed.ai or runware.ai for current rates.

Resolution lists every size a model can output; HD means it has no resolution setting at all and always returns roughly 1024×1024. Prices are per image, charged straight to whichever provider you’ve connected. The figures here are WaveSpeed’s rates; Runware prices its own catalogue. Where there’s a range, the low end is the smallest size at the lowest quality and the high end is 4K at the highest, so the setting you pick matters as much as the model you pick.

More advanced models can take quite a few reference images at once, but the count matters far less than the choice. References pulling in different directions give the model contradictory instructions, and you get muddle instead of the result you were after — so add an image because it says something the others don’t, not to fill the slots.

Switching to a different model resets your settings, so nothing carries over by mistake. And if you connect more reference images than a model can use, the extra ones just get greyed out rather than removed — switch to a model that supports more, and they come back to life.

← Back to all articlesNeed help? Go to Support →