Image Generation

Learn how to choose image models and settings, generate images, and compare multiple versions in CawCut workflows.

2026년 8월 31일 업데이트 · 12분 읽기

Image Generation nodes let you create new visuals from text or transform existing images inside a CawCut workflow. You can select the model and output settings that fit your task, connect reference images when needed, and generate several versions in one run to find the result you want.

1. Image Input Nodes

Use an Image Node or Image Group Node to add source and reference images to a workflow before connecting them to an Image Generation node or another compatible node.

1.1 Image Node

The Image Node adds one image to the workflow. Drag and drop an image into the node, select [Upload] to add one from your device, or select [Select Asset] to use an existing CawCut asset. Connect the Image output to another node to use the uploaded image as an input.

To add annotations or masks that more precisely guide downstream AI models, see Image Editing and Enhancement Nodes.

An Image Node with controls for uploading an image or selecting an existing asset.

1.2 Image Group Node

The Image Group Node keeps multiple images together so they can be passed through the workflow as one image group. Open the add node menu, select [Group], then select [Image Group] to add it to the canvas.

Upload images or select existing assets, then use the checkmark on each thumbnail to choose which images are included in the output. The output label shows the number of selected images.

An Image Group Node containing three images with two selected for its output.

2. Image Generation Node

The image generation node.

The Image Generation node supports two generation modes in the same Node Panel:

  • Text to Image: Describe the image you want with a prompt.
  • Image to Image: Add one or more images, then explain how CawCut should transform, combine, or use them.

The node also includes:

  • A Prompt input for text supplied by another node.
  • Image input ports for reference or source images. The node can display inputs for up to 16 images, although individual model limits may vary.
  • An Image output that passes the selected result to connected nodes.
  • Controls for the model, aspect ratio, resolution, quality, and number of generated images.
  • An estimated credit amount beside Run, which updates according to the selected generation options.

3. Configure Node Settings

Use the controls along the bottom of the Node Panel before selecting Run. These settings determine how CawCut generates the result and can affect its dimensions, level of detail, generation time, and credit use.

3.1 Choose an Image Model

Open the model menu and choose the model that best matches the image you need.

ModelBest for
GPT Image 2Complex general-purpose generation and editing, strong prompt following, text, and detailed reference-based work
GPT Image 1Reliable general image generation and editing across a wide range of styles
GPT Image 1 miniFaster, lower-cost drafts and straightforward image tasks
Nano Banana 2Fast high-volume creation, conversational editing, multiple references, consistency, and readable text
Nano Banana ProComplex professional assets, precise creative control, brand consistency, product mockups, and grounded visuals
Nano BananaFast general-purpose generation and straightforward reference-based edits
Flux 2 ProProduction-quality generation and editing, photorealistic work, and multi-reference compositions
Seedream 5.0 ProPrecise editing, multilingual text, product visuals, marketing assets, and professional production
Seedream 5.0 liteFaster all-purpose generation, image series, and reference-heavy workflows
Seedream 4.5High-resolution generation, multi-image work, and workflows already designed around Seedream 4.5
Ideogram 4.0Typography, posters, controlled layouts, multilingual text, and design-focused images
Z Image TurboFast photorealistic drafts, high-volume iteration, and Chinese or English text
Recraft V3Graphic design, vector-style artwork, icons, brand visuals, mockups, and text-heavy creative work
Wan 3.6Image-to-image transformations, creative variations, and prompt-guided edits

3.2 Model Support by Generation Mode

Use this table to quickly compare which current models are available for each image generation mode.

ModelsText to ImageImage to Image
GPT Image 2 / GPT Image 1 / GPT Image 1 mini
Nano Banana 2 / Nano Banana Pro / Nano Banana
Seedream 5.0 Pro / Seedream 5.0 lite / Seedream 4.5
Ideogram 4.0
Flux 2 Pro / Z Image Turbo / Recraft V3
Wan 3.6

3.3 Model Specifications

The table shows models whose current data includes at least one documented limit. Prompt length is measured in characters. The Reference images column shows how many image inputs the model can use in Image to Image mode. The Images per run column shows the selectable output range when the current model data defines one.

ModelText to Image promptImage to Image promptReference imagesImages per run
GPT Image 21–50,0003–50,0001–161–4
GPT Image 11–4,0003–50,0001–41–4
Nano Banana 23–1,0001–2,0001–141–4
Nano Banana1–4,0001–4,0001–41–4
Seedream 5.0 Pro3–20,0003–20,0001–10Not supported
Seedream 5.0 liteNot listed3–50,0001–41–4 for Image to Image
Seedream 4.53–10,0003–10,0001–4Not supported
Recraft V31–2,0001–4 for Text to Image
Wan 3.6Up to 50,0001–31–4 for Image to Image

The current model data does not list prompt or reference-image limits for every available model. When a value is not listed, use the validation message and available controls shown in the Node Panel.

3.4 Choose an Aspect Ratio

The aspect ratio controls the shape of the generated image. Open the ratio menu and choose the format that matches the final destination.

Common choices include:

  • 16:9 for landscape image, presentations, and wide banners.
  • 9:16 for vertical image and mobile-first content.
  • 1:1 for square social posts and profile graphics.
  • 21:9 for extra-wide cinematic images.
  • 3:2, 4:3, and 5:4 for common landscape images.
  • 2:3, 3:4, and 4:5 for portrait images and posts.

Choosing the final ratio before generation usually gives the model more room to compose the scene correctly than generating first and cropping later.

3.5 Choose a Resolution

Resolution controls the output size in pixels. Depending on the selected model, CawCut may offer 1K, 2K, or 4K.

ResolutionBest for
1KPrompt tests, drafts, quick comparisons, and smaller online uses
2KGeneral final images, larger social content, and moderate cropping
4KLarge displays, detailed final assets, print-oriented work, and close crops

Higher resolution creates a larger file and may take more time or credits. It does not automatically fix prompt, composition, anatomy, or text problems, so confirm the overall result at a lower resolution before using a larger final setting when possible.

3.6 Choose a Quality Level

Quality controls how much generation effort the selected model applies to the result. Open the quality menu and choose Low, Medium, or High.

QualityBest for
LowFast drafts, early prompt testing, and broad composition ideas
MediumA balance of speed and detail for most everyday generations
HighFinal results where refined textures, edges, text, or small details matter

The comparison below uses GPT Image 2, 1K resolution, and the same overall task at Low, Medium, and High quality. The enlarged areas make small surface details easier to inspect.

Low, Medium, and High quality GPT Image 2 results with enlarged detail areas.

In the comparison above, Low is noisier and misses some details, Medium preserves more detail but still contains some noise, and High removes most of the noise and produces more realistic detail.

4. Generate Multiple Versions

Generating multiple versions in one run gives you several interpretations of the same prompt and inputs. You can compare composition, details, and overall style, then keep the result that best matches what you want.

4.1 Choose the Number of Images

Open the image amount control at the bottom of the Image Generation node and choose how many images you want to create. Then select Run.

The number of images control in an Image Generation node, highlighted with two images selected.

All versions in the run use the same prompt, image inputs, model, aspect ratio, resolution, and quality settings.

4.2 Quickly Select a Version

After generation is complete, open the versions menu on the image. Select the version you prefer, then select Confirm. The selected version becomes the node's active image and is passed to any connected nodes.

The versions menu showing four generated images and the control for selecting a version.

4.3 Compare Versions in the Full Preview

If you need a closer look, select the [Preview] icon on a version to open it at a larger size. Use the thumbnails on the left to compare the generated results, then select Use this version when you find the best one.

관련 기사

Frequently asked questions

Learn how Image Generation nodes, models, settings, references, and multiple versions work in CawCut.

The Image Generation node can create an image from a text prompt or transform existing images with a prompt. You can choose a model, aspect ratio, resolution, quality level, and the number of versions to generate.

The currently available choices include GPT Image 2, GPT Image 1, GPT Image 1 mini, Nano Banana 2, Nano Banana Pro, Nano Banana, Flux 2 Pro, Seedream 5.0 Pro, Seedream 5.0 lite, Seedream 4.5, Ideogram 4.0, Z Image Turbo, Recraft V3, and Wan 2.6. Available models depend on whether you select Text to Image or Image to Image and may change as CawCut adds or updates models.

Choose the model based on the task. For example, use Ideogram 4.0 for typography and controlled layouts, Recraft V3 for design-oriented graphics, Z Image Turbo for fast drafts, or a premium general-purpose model such as GPT Image 2 or Nano Banana Pro for complex generation and editing. If you are unsure, test the same prompt with a few models.

Text to Image creates a new visual from your prompt. Image to Image uses one or more images as references or source material and follows your prompt to transform, combine, or restyle them.

The Image Generation node can display inputs for up to 16 images, but the usable number depends on the selected model. The current documented maximum ranges from 3 reference images for Wan 2.6 to 16 for GPT Image 2.

Match the ratio to where the image will be used. Common choices include 16:9 for landscape video, 9:16 for vertical video, and 1:1 for square posts. CawCut also offers additional portrait and landscape ratios when supported by the model.

Higher resolutions create larger images with more pixels and can preserve more detail for large displays, close crops, or final delivery. 1K is usually enough for quick tests, while 2K or 4K is better when the selected model supports it and you need a larger final image.

Low is useful for faster drafts, Medium balances speed and detail, and High is intended for more refined final results. Higher quality can require more generation time or credits, and the visible difference depends on the model, prompt, resolution, and image content.

Open the image amount control, choose how many images you want, and select Run. CawCut generates the selected number of versions using the same prompt, inputs, model, and settings.

Change one factor at a time, such as the prompt, reference image, model, aspect ratio, resolution, or quality. Generate multiple versions after each focused change so you can identify which adjustment improves the result.