AI image model
Grok Imagine AI Image Generator
Choose Grok Imagine when one source image is enough to set the visual starting point and the Standard or High mode should be selected before generating.
Grok Imagine task directions
These local references help turn the current Grok Imagine input and control contract into a concrete task direction. They are not presented as output from this model.
Every card identifies whether it is a verified model-family example or a visual reference. Reference cards are not model benchmarks or before-and-after claims.
- Input contract
- Text prompt + up to 1 reference image
- Current controls
- Prompt length: Up to 5,000 characters · Output behavior: Provider-selected output (no resolution selector) · Framing: 5 aspect-ratio choices
- Selection boundary
- Choose a multi-reference model when details from several source images must inform one result. Choose a model with manual 1K, 2K, or 4K tiers when the delivery size must be selected in advance.

Subject and setting
Use one clear subject and setting to test whether the written brief establishes the right visual starting point.
- Task direction
- Use one clear subject and setting to test whether the written brief establishes the right visual starting point.
Reference direction, not a model output
Local ImageStyle visual reference. It is not presented as a model-specific result.

Composition direction
Use a deliberate frame and hierarchy to decide what the prompt or reference must preserve.
- Task direction
- Use a deliberate frame and hierarchy to decide what the prompt or reference must preserve.
Reference direction, not a model output
Local ImageStyle visual reference. It is not presented as a model-specific result.

Detail and material cue
Specify the visible detail that should guide the result instead of treating the reference as a performance test.
- Task direction
- Specify the visible detail that should guide the result instead of treating the reference as a performance test.
Reference direction, not a model output
Local ImageStyle visual reference. It is not presented as a model-specific result.
- Create
- AI images
- Input
- Text prompt + up to 1 reference image
- Output behavior
- Provider-selected output (no resolution selector)
Compare available settings with related models
Check whether the input, output, and framing controls fit the job before you open a workspace. This table does not treat the example media as a cross-model performance test.
Grok Imagine
Current page
A text or single-reference image direction with Standard or High mode selection.
- Input
- Text prompt + up to 1 reference image
- Prompt length
- Up to 5,000 characters
- Output behavior
- Provider-selected output (no resolution selector)
Grok Imagine Image 2.0
Text-to-image with a 5,000-character prompt limit and five documented frames.
- Input
- Text prompt only
- Prompt length
- Up to 5,000 characters
- Output resolution
- 1K
Nano Banana 2
Reference-guided images with manual 1K, 2K, or 4K selection and broad framing.
- Input
- Text prompt + up to 5 reference images
- Prompt length
- Up to 20,000 characters
- Output resolution
- 1K / 2K / 4K
Nano Banana 2 Lite
Reference-guided image directions with broad framing and provider-selected output.
- Input
- Text prompt + up to 5 reference images
- Prompt length
- Up to 20,000 characters
- Output behavior
- Provider-selected output (no resolution selector)
What you can control with Grok Imagine
Use one image as a deliberate starting point
Grok Imagine currently accepts one reference image and exposes Standard and High modes. Keep the prompt focused on the change or composition that the source image should guide.
Add reference images when they help
Grok Imagine currently supports up to 1 reference image. Only upload source material you are allowed to use.
Confirm the current output behavior
Current output tiers are Provider-selected output (no resolution selector), with 5 aspect-ratio choices. Use the options shown in the workspace before generation.
Current Grok Imagine settings
These settings come from the current ImageStyle model configuration and help you decide whether the workflow fits the job.
| Setting | Current availability |
|---|---|
| Input | Text prompt + up to 1 reference image |
| Prompt length | Up to 5,000 characters |
| Output behavior | Provider-selected output (no resolution selector) |
| Framing | 5 aspect-ratio choices |
AI image model
When to choose Grok Imagine
A text or single-reference image direction with Standard or High mode selection.
Current workflow fit
Choose Grok Imagine when one source image is enough to set the visual starting point and the Standard or High mode should be selected before generating.
When to compare another model
Choose a multi-reference model when details from several source images must inform one result. Choose a model with manual 1K, 2K, or 4K tiers when the delivery size must be selected in advance.
How to create an image with Grok Imagine
Decide whether the prompt alone or one approved image sets the visual direction, then choose the current Standard or High mode and frame.
- 1
Choose text-only or one-image direction
Use the source image only when it establishes the subject or composition; otherwise write a full prompt and select the frame and mode.
- 2
Add references when needed
Upload approved reference images only when a defined visual starting point will help the result.
- 3
Confirm the framing
The model selects the output tier; confirm 5 aspect-ratio choices before generation.
- 4
Generate and review
Before public use, inspect visible text, product details, people, marks, and framing in the result.
Checks before you create and publish
- Treat the workspace as current
- Model availability and settings can change. Use the options shown in the current workspace before you submit.
- Use approved source material
- When you upload a reference image, confirm that you are allowed to use and publish it for the current project.
- Review the generated result
- Manually check any visible text, product information, and brand element that must be accurate before publishing.
Grok Imagine FAQ
- What are Grok Imagine image editing capabilities in ImageStyle?
- The current workflow accepts one approved reference image, a written edit direction, five framing choices, and Standard or High mode. Review text, faces, and product details before publishing the result.
- Can Grok Imagine combine multiple reference images?
- No. The current ImageStyle configuration accepts one reference image. Use a multi-reference image model when several source images need to guide the result.
- Can I use reference images with Grok Imagine?
- Yes. The current workflow supports up to 1 reference image. Only upload images you are allowed to use.
- What output settings does Grok Imagine currently offer?
- Current output tiers are Provider-selected output (no resolution selector), with 5 aspect-ratio choices. The generation workspace is the source of truth for the options available at submission.
- What should I prepare before using Grok Imagine?
- Prepare a clear visual direction. If the current model supports reference images, add approved images only when they help preserve a visual starting point.
- Can I publish a generated result without review?
- No. Before public use, manually review text, product details, people, marks, and any other element that must be accurate.
Updated:
Capability source: Current ImageStyle image-model configuration.