FREE image generation with no sign-up, plus daily check-in credits that climb all week. No credit card, ever.

Grok Imagine Image 2.0 Generator

Grok Imagine Image 2.0 turns your prompt and up to five references into posters, product shots and candid photos with text that reads, in 3:2 to 9:16 up to 2K.

Text you can readFive references, one imageNo weekly limit

Grok Imagine Image 2.0 Generator

45 creditsSign in to run this one
Examples

Grok Imagine Image 2.0 images creators posted

Each posted on X by its maker, with the model named.

Wildlife that looks shot on location
A manga page lettered in Japanese
A product ad, label and all

Wildlife that looks shot on location

How creators work

Build an image with Grok Imagine Image 2.0 in four moves

The habits that show up again and again in the prompts people share.

01

Pick the frame shape first

Square, 16:9, 9:16, 3:2 or 2:3, at 1K or 2K.

02

Write down everything that must be there

Wording, layout, where the hands go. It renders what you name.

03

Add up to five references

Switch to image-to-image and upload the face, the outfit, the product.

04

Turn the still into a video

Use it as the first frame in Seedance 2.5, Kling 3.0 or MiniMax H3, all on the same balance.

Access

Grok Imagine Image 2.0 without SuperGrok, Replicate or an API key

No weekly quota and nothing to install. You pay credits per image, see the price before you press, and download the file.

Grok Imagine Image 2.0 — a continental breakfast table
A continental breakfast laid on a lace cloth: croissants, sliced bread, butter, jam, coffee, orange juice and a bowl of berries. Posted by @karatademada as made with Grok Imagine Image 2.0.

Creative engine

Grok Imagine Image 2.0 on Kavel

5
Reference images per run
2K
Top resolution
45 credits
Per 1K image

Up to five reference images

Upload them in image-to-image mode: a face, an outfit, a product, a colour palette.

Print-shaped 3:2 and 2:3

For posters, photo prints and book covers, at 1K or 2K.

The price before you press

45 credits at 1K, 60 at 2K, the same when you upload references.

What people post about

Three things Grok Imagine Image 2.0 does best, according to creators

Text that reads, from infographics to handwritten pages

Text that reads, from infographics to handwritten pages

This whole page came out of one generation: numbered sections, boxed diagrams and paragraphs of hand lettering. Creators on X post infographics and Japanese manga built the same way.

  • Put every word in quotes and say where it sits
  • Name the type: oversized serif, handwritten marker, bold sans
  • Split a busy page into numbered sections
Candid photos, when you prompt them like a snapshot

Candid photos, when you prompt them like a snapshot

A phone selfie against a white door in soft window light, no studio polish. On r/grok the prompts that look most real read like camera notes.

  • Describe a simple scene: who, where, what light
  • Add snapshot cues: candid, amateur, old iPhone
  • Say where the hands are and what they hold
Web heroes, posters and brand layouts

Web heroes, posters and brand layouts

A furniture landing page with a serif headline, nav bar and floating content cards, from one template prompt. Designers on X share web, poster and 3D icon templates built on it.

  • Name the layout: headline left, product right
  • List the palette as named colours
  • Keep the template and swap only the [SUBJECT] line
Prompts

Grok Imagine Image 2.0 prompts you can copy

Creators' own prompts, word for word. The folk poster is a template: replace [SUBJECT].

Low-angle selfie under a summer sky

Low-angle selfie under a summer sky

@woleswoosh

Written for 9:16, 2K, medium quality: the settings this model runs at here.

A low-angle selfie portrait of a young Southeast Asian woman with fair skin and soft features, looking slightly upward at the camera with a playful, pouty expression and slightly parted lips. She has large, expressive dark brown eyes with long, thick black eyelashes and subtle makeup (soft pink blush, natural brows, glossy pink lips). She is wearing clear transparent rectangular eyeglasses with thin frames. Her hair is completely covered by a soft, light beige / nude-toned hijab (headscarf) draped loosely around her head and neck in natural folds, with a smooth, slightly textured fabric that catches the light.

She is raising her right hand near the left side of the frame in a classic peace / V-sign gesture (index and middle fingers extended upward, other fingers folded), with neatly manicured short white nails. She is wearing a light gray long-sleeved top with subtle ribbed or textured fabric details on the sleeves and chest, and layered silver chain necklaces — one thicker curb-style chain and a thinner chain with a small rectangular metal pendant hanging down the center of her chest.

The background is a bright, clear blue sky filled with fluffy white cumulus clouds, shot from a low angle so the sky dominates the upper portion of the frame. Soft natural daylight illuminates her face from above and slightly to the side, creating gentle highlights on her cheekbones, the lenses of her glasses, and the folds of the hijab. The overall mood is cheerful, casual, and youthful; photorealistic style with sharp facial details, natural skin texture, and realistic fabric folds. Vertical composition, close-up framing focusing on her face, hand, and upper torso.
Run this prompt
Folk screenprint poster (template)

Folk screenprint poster (template)

@doganuraldesign

Earthy palette, oversized lettering and botanical borders around your subject.

Illustration of [SUBJECT], flat folk graphic illustration, hand-drawn vector forms, screenprint aesthetic, limited earthy palette, deep forest green, warm mustard, cream, burnt orange, muted navy, teal accents, bold oversized English typography, ornamental botanical motifs, geometric pattern blocks, layered collage composition, imperfect ink registration, distressed print texture, subtle paper grain, rough-edged shapes, playful asymmetry, handcrafted editorial poster feel, warm nostalgic atmosphere
Run this prompt
What people say

What creators said after trying it

Posted on X in launch week, quoted as written.

“

Grok Imagine Image 2.0 is amazing with text rendering!

Compared

Grok Imagine Image 2.0 vs GPT Image 2.5, GPT Image 2 and Nano Banana

The matchups people ran after launch.

vs GPT Image 2.5

One reference photo through three models on r/generativeAI. Pick Grok for print-shaped 3:2 and 2:3 frames built from up to five references.

“Grok seems to be more cute and straight tot he point.”

See GPT Image 2.5

vs GPT Image 2

Posted side by side in launch week as the two leaders on Arena. Grok takes up to five references in a single run.

“literally the top 2 image models as per Arena side by side”

See GPT Image 2

vs Nano Banana 2

A Japanese pixel artist added a grid to his prompt and got true pixel art with exact colour codes, something he said neither ChatGPT nor Nano Banana had managed.

See Nano Banana 2

vs Nano Banana Pro

A r/grok user who runs reference edits through many APIs put only GPT Image 2 and Nano Banana Pro in the same league for working from references.

See Nano Banana Pro
FAQ

Grok Imagine Image 2.0 FAQ

The questions people ask most in Grok Imagine threads.

You can start free: log in and Kavel adds free credits with no card, which run Kavel's own image engine and Nano Banana 2 Lite. Grok Imagine Image 2.0 runs on any paid plan or credit pack, cost shown before you generate.

No. You spend credits per image, 45 at 1K and 60 at 2K, and nothing resets on a weekly clock. Failed runs are refunded automatically, unless the prompt broke the content policy.

Keep the prompt simple and write it like a snapshot: who, where, what light. Add cues such as "amateur, candid, old iPhone" and skip boosters like photorealistic, ultra-detailed or gorgeous.

Name the finish you want. r/grok users get it with "flat colouring anime screen cap" or "1990s OVA anime" at the start of the prompt.

Up to five per run. Switch to image-to-image, upload them, and say what each one is for: "the face from image 1, the jacket from image 2".

Yes. Creators posted full manga pages with Japanese dialogue in every balloon. Quote the exact words in your prompt and say which panel or balloon they go in.

Yes. Generate the character, set or product here, then open the video generator and use it as the first frame in Seedance 2.5, Kling 3.0 or MiniMax H3, on the same credit balance.

No. Pick the model, type a prompt or upload references, generate in the browser and download the image.

Make your first Grok Imagine Image 2.0 image

The poster, the product shot, the page full of words: generate it with readable text and up to five references, from 45 credits at 1K with no weekly quota, and if a run errors on our side the credits go straight back.