Grok Image 2.0: Imagine Image 2.0 Generator

Plan an Imagine Image 2.0 task from text or one to three reference images, with visible resolution, quality, output-count, format, and credit controls.

Model
Imagine Image 2.0
Inputs
Text · 1–3 references
Settings
1K / 2K · 1–20 outputs

Loading generator…

What Is Grok Image 2.0? Meet Imagine Image 2.0

Editorial planning board for a controlled image-edit brief with source, change, and review notes
Planning artwork separates source, intended change, and review note before an image task. It is not an Imagine Image 2.0 output, interface, task record, or delivery evidence.

On ZMS AI, Grok Image 2.0 is a search phrase for Imagine Image 2.0, not a family page. Its workspace identifier is grok-imagine-image-2.0.

Prepare text-to-image or image-to-image with one to three references. The contract and estimate are visible, but a paid signed-in test must prove submission, returned image URLs, and final debit. This page presents planning material, not results.

Choose a Text-to-Image or Image-to-Image Path

Start with your material. The ZMS AI image workspace covers broader workflows; this page holds Imagine Image 2.0 task choices.

Planning rail that separates text, one reference image, and multiple reference-image roles
Planning rail separates text, one-reference, and multi-reference roles. It is not a live control, upload screen, or Imagine Image 2.0 output.
PathPrepareConfirmed boundary
Text-to-imageA written brief and one ratio pair.26 dimensions: 13 ratios at 1K or 2K.
Image-to-imageOne to three references and a change request.Reference editing is 1K/2K without dimension selection.
Shared selection rulesReview quality, count, format, and estimate.Both paths allow low/medium quality and 1–20 outputs.

Use Only the Confirmed Image Settings

These cards document the selector boundary; they do not imply hidden controls.

13 ratio pairs

Text-to-image offers 13 ratios, each at 1K and 2K.

Reference resolution

With references, choose 1K or 2K; image-to-image removes ratio selection.

Low or medium

Quality is low or medium; a production task still needs review.

One to 20 outputs

Choose 1–20 outputs. The estimate is not final debit.

The 26 Text-to-Image Dimensions

Choose one row and one resolution for a text-only task. Reference-led edits show only their 1K or 2K resolution selector instead of these width and height pairs.

Aspect ratio1K dimensions2K dimensions
1:11024 × 10242048 × 2048
3:4864 × 11521776 × 2368
4:31152 × 8642368 × 1776
9:16720 × 12801584 × 2816
16:91280 × 7202816 × 1584
2:3832 × 12481664 × 2496
3:21248 × 8322496 × 1664
9:19.5576 × 12481344 × 2912
19.5:91248 × 5762912 × 1344
9:20576 × 12801440 × 3200
20:91280 × 5763200 × 1440
1:2704 × 14081456 × 2912
2:11408 × 7042912 × 1456

See the Imagine Image 2.0 Credit Estimate

Each row is a per-output estimate. Multiply it by 1–20 outputs, then use ZMS AI credit packs when ready to fund a task; final debit needs E2E verification.

ResolutionQualityCredits per output
1KLow12 credits
1KMedium18 credits
2KLow18 credits
2KMedium24 credits

Start with Suggested Text-to-Image Prompts

These are editable planning inputs, not images or samples. For writing guidance, read the GPT Image 2 prompt guide, then select settings here.

Four text-only starting briefs

4 starter recipes
Text to Image

Light a deliberate still life

Define object, set, and material.

Canvas
4:3 · 1K
Quality
Low
Outputs
1
Format
JPG
Suggested prompt

Studio still life: one brushed steel travel mug on a limestone plinth, three-quarter view, daylight from upper left, grey paper backdrop. Keep one handle, clean rim, realistic highlights, small shadow; no label, extra objects, hands, liquid, or glare.

Review checkpoint

Check handle count, rim, reflections, shadow, and type.

Text to Image

Leave room for poster type

Reserve a text-safe field without lettering.

Canvas
3:4 · 1K
Quality
Medium
Outputs
2
Format
JPG
Suggested prompt

Vertical poster background: a cobalt paper wave rises from the lower third on bone white. Keep the upper half open for typesetting, with one shadow, crisp edges, subtle grain, and no words, logos, symbols, people, or extra objects.

Review checkpoint

Check upper-field space and absent lettering.

Text to Image

Direct an editorial portrait

Specify light, posture, wardrobe, and frame.

Canvas
2:3 · 2K
Quality
Medium
Outputs
1
Format
JPG
Suggested prompt

Editorial half-length portrait of an adult ceramic artist, daylight studio, looking just past camera, charcoal shirt, neutral apron, hands at waist. Soft right window light, clay palette, 85mm perspective; no text, logos, extra hands, duplicated tools, distorted shelves, or retouching.

Review checkpoint

Check hands, shelf geometry, eye direction, and light.

Text to Image

Frame an environmental beat

Anchor location, time, and path.

Canvas
16:9 · 2K
Quality
Low
Outputs
3
Format
JPG
Suggested prompt

Wide evening scene of one cyclist in a coastal town, viewed from behind. A narrow road curves to low white buildings, sea beyond the roofline, coral sky fading blue. Keep one bicycle and natural perspective; no crowd, signage, text, extra bicycles, impossible architecture, or flare.

Review checkpoint

Check bicycle count, road, horizon, scale, and sky.

Write a Reference-Led Edit Brief

When an image-to-image task needs one to three references, write down what each source may influence before you upload it. The useful boundary is not a hidden editing tool; it is a brief that makes the result easier to review after an actual task returns, with clear source roles and review criteria.

Planning board separating source roles, permitted changes, materials, and review checks for an image edit
ZMS planning artwork assigns source roles, permitted changes, material constraints, and review checks. It is not generated output, live UI, or proof of completed delivery.
  1. 01

    Assign each source role

    Name each source as composition, material cue, or subject constraint; keep every reference role distinct.

  2. 02

    Write what must stay

    List identity, camera position, object count, geometry, or light direction that must stay clearly recognizable.

  3. 03

    Limit permitted changes

    Describe swap, cleanup, mood, or material adjustment without inventing masks, brushes, search, or seed controls.

  4. 04

    Plan a human review

    List returned-image details carefully to compare so a plausible result is never accepted without inspection.

Suggested edit brief

Use the first uploaded image only as the composition and camera reference: retain the seated adult subject, three-quarter crop, tabletop position, window direction, and visible hand count. Use the second uploaded image only as a material reference for the dark green glazed ceramic cup; keep its subtle speckle and satin reflection, but do not copy logos or lettering. Change the room from a bright daytime café to a quiet early-evening study with a deep blue wall and one warm practical lamp behind the subject. Preserve realistic table perspective, natural fingers, a single cup, and a readable boundary between lamp glow and window shadow. Do not add people, signs, text, duplicate cups, altered facial features, extra hands, dramatic haze, or a different camera angle.

  • Confirm subject, crop, hand count, perspective, and cup count.
  • Check ceramic cue without copied lettering or unrelated objects.
  • Review lamp direction, stable shadow, and scene geometry.

What a Signed-In E2E Test Still Needs to Prove

Only a signed-in paid E2E run can connect the request, returned media, and debit.

E2E checkWhat the test must showWhat this page does instead
Task creationSigned-in submission accepts the selected request and visible settings.Keep Generate and validation without delivery claims.
Status and image URLsPolling returns task state and image URLs in the library.Publish no results, speed, quality, or reliability claims.
Final debitDebit matches selected resolution, quality, and output count.Show an estimate and retain noindex, follow.

Imagine Image 2.0 and Grok Image FAQ

These answers separate the search phrase, the specific model workspace, the current request boundary, and the E2E evidence that is still pending.

Is Grok Image 2.0 the official model name?

No. Grok Image 2.0 is the search phrase served by this page. The specific workspace name is Imagine Image 2.0, and the current model identifier is grok-imagine-image-2.0.

Is this a Grok Imagine family page?

No. This page is limited to the Imagine Image 2.0 version workspace. A future /grok-image route can cover family-level choices without duplicating this model-specific contract.

How many text-to-image dimensions are shown?

Text-to-image exposes 13 aspect-ratio rows, each with one 1K and one 2K width-by-height pair. That makes 26 confirmed text dimension choices.

How many references can an image-to-image task use?

Use one to three reference images. In that reference-led path, the workspace exposes 1K or 2K resolution rather than a separate width and height selector.

Which quality and output-count settings are visible?

The current workspace exposes low or medium quality and one through 20 requested outputs. The estimate multiplies the selected per-output rate; final debit still needs E2E evidence.

Why are there no result examples or verified availability claims?

A signed-in paid task must still prove creation, returned image URLs, and final debit together. Until then, this page does not present a Showcase, performance conclusion, reliability claim, or verified delivery statement.