On August 7, 2026, SpaceXAI released Imagine Image 2.0. It is live now as the new Quality Mode inside Grok, on the web and in the iOS and Android apps.
Most image model launches sell you a picture. This one sells you a process.
Read the launch page and you will notice something odd. The demo is not really about beauty. It is about editing one region and leaving the rest alone. It is about cutting a subject out with a clean edge. It is about resizing one asset into nine frames. It is about fifteen ready-made templates, four of which are filed under "Marketing."
That is not a gallery. That is a studio's to-do list.
We have written about image models before, including Google putting image generation inside Search. This one is different in kind, not degree. Here is what shipped, where it actually ranks, and the one catch that decides whether you can use it this week.
What SpaceXAI actually shipped
Imagine Image 2.0 is generally available. There is no waitlist and no beta flag. You open grok.com/imagine, pick Quality Mode, and you are in.
The company frames the goal plainly. It says it built the model to "make images you can use in real work." It also claims the model plans typography and layout the way a designer would. The point of that is dense, multi-part visuals that hold together, with small text that stays sharp.
Five things are new. Each one maps to a job a human currently does by hand.
Magic wand. You point at a region. The model changes only that region. The rest of the frame stays as it was.
Segmentation. You select a precise area to change. This is the difference between a rewrite and an edit.
Background removal. Any subject exports with a transparent background. It drops straight into other work.
Multi-reference editing. One generation accepts up to five input images. That removes the manual compositing step that used to mean stitching sources together yourself.
Smart resize. You pick a ratio and the model fills in the frame. Nine ratios ship at launch: 1:2, 9:16, 2:3, 3:4, 1:1, 4:3, 3:2, 16:9 and 2:1.
Look at that list again. None of it is about having an idea. All of it is about finishing an asset.

The templates are the real story
SpaceXAI also shipped templates. Each one packages a common workflow into a starting point, with the steps already configured. You supply the inputs and get a finished result.
Fifteen templates went live across six categories. Here is the full set, grouped the way the launch page groups them.
- Photo Tools: Photo Edit, Reimagine, Photo Collage, BG Removal and Change, Professional Headshot
- Product: Product Color Change
- Marketing: Editorial Product Poster, E-Commerce Photos, UGC Photos, Merch Maker
- Design Tools: Mascot Maker, Icon Maker, Character Sprite
- Game Assets: Props and UI Kit
- Streaming: Emoji Creator
Now read the four filed under Marketing again. Editorial Product Poster. E-Commerce Photos. UGC Photos. Merch Maker.
Those are not features. Those are line items on an agency invoice.
For two years, image AI competed on the first ten minutes of a job. Can it make something that looks good? That question is mostly settled. The launch page has moved to the next eight hours. Can it version, resize, cut out, composite and finish?
That is the shift worth naming. The moat was never the idea. It was the production line.
Quick Facts: Imagine Image 2.0 at a Glance
- Released August 7, 2026, generally available as Quality Mode — (Source: SpaceXAI, 2026 — official launch post)
- Ranks second in both the Arena image-edit and text-to-image boards — (Source: SpaceXAI, 2026 — official launch post)
- Elo of 1,439 in image editing against GPT-Image-2 at 1,463 — (Source: The Decoder, Aug 8 2026 — Arena benchmark report)
- Accepts up to five reference images in one generation — (Source: SpaceXAI, 2026 — official launch post)
- API access is not available yet — (Source: SpaceXAI, 2026 — official launch post)
Where it ranks, and the detail most coverage skipped
SpaceXAI says the model ranks second in the world in both text-to-image generation and image editing. It cites the Arena leaderboards as of August 7, 2026.
Second is the honest word here. Not first.
OpenAI's GPT-Image-2 leads both boards. In image editing it scores an Elo of 1,463 against 1,439, as The Decoder reported on August 8. That is a gap of 24 points. In text-to-image the gap is wider: 1,380 against 1,320, or 60 points.
So the headline is real, and the margin is not close in one of the two cases.
Now the detail almost nobody mentioned. Look at the label on the ranked entry. It reads "grok-imagine-image-2 (low)". The low setting is what placed second, on both boards. A separate SpaceXAI entry, listed as "grok-imagine-image-quality", sits further down each board.
One more quirk to file away. On Arena, the company's models are listed under the name SpaceXAI, not xAI. If you go looking for the row yourself, look for that.
Also worth remembering: an Arena score is a human preference vote, taken as a snapshot on one day. It is not a measure of whether the model can hold a brand style across forty assets. Only your own test tells you that.

The catch: there is no API
Here is the line that decides your week. SpaceXAI says API access is "coming soon". It is not here.
That inverts the usual advice. Normally a launch like this ends with "wire it into your pipeline". You cannot. Not yet.
What you get instead is a consumer product. A browser tab and two phone apps. So the value today is manual and it lands with your designers, not your engineers.
That is not nothing. It is arguably better for a small team. Nobody has to build anything. But it does mean three things.
You cannot batch. Four hundred product shots stay a human sitting at a screen.
You cannot version-control it. There is no prompt log in your repo and no reproducible run.
You cannot cost it per asset. Until pricing exists, you are estimating.
If your plan depended on an API, park the plan. Put a reminder in your calendar instead.
What changes inside a studio week
Think about where hours actually go on a creative job.
Very few go to the first idea. Most go to what comes after. Resize this for the story frame. Cut the model out. Swap the product colour for the second SKU. Make a version for the poster and a version for the banner. Match last month's look.
Every new tool on this launch page attacks that second pile.
Multi-reference editing is the clearest example. If you can feed five images into one generation, you are describing a composite. A product, a background, a model, a texture, a reference for lighting. That used to be a layered file and an afternoon.
Smart resize is the second. Nine ratios from one asset is a social team's entire delivery list.
Background removal is the third, and it is the least glamorous and most useful. Clean cutouts are a tax every e-commerce team pays every month.
We made a similar point when Alibaba's Qwen model was graded on running a shop for a year. The benchmarks are moving from answering to operating. This launch is the same move, in pictures.
How it sits next to what you already use
Nobody starts from zero here. Most teams already run some mix of tools.
If you use Photoshop and its generative fill, the overlap is the region edit. The difference is where the work happens. One is a layered file you own. The other is a browser tab you do not.
If you use a dedicated cutout tool, the overlap is background removal. Test them head to head before you switch. Edge quality on hair, glass and fabric is where these tools usually fail.
If you use another image model for concepts, the overlap is smaller than you think. This launch is not really aimed at the concept stage. It is aimed at everything after it.
And if you use a stock library, the pressure point is the template list. E-Commerce Photos and UGC Photos target the exact briefs teams currently solve by licensing an image. That is a budget line worth revisiting.
The sensible read is not "replace the stack". It is "find the one step in your stack that eats the most hours". Then test this against that step only. A single honest comparison beats a full migration built on a launch post.

The consistency question
There is one more section on the launch page that is easy to scroll past.
It is called "Build a world for video". The demo generates a character, her locations and her props as separate images, and holds one style across all of them.
Consistency is the thing brands actually need. Not one great image. Forty images that look related.
Every brand team that has tried AI imagery has hit this wall. The hero shot is lovely. The next nine look like a different company made them. That is why most AI imagery still lives in moodboards and never reaches a campaign.
If holding a look across generations really works, that is the feature that matters more than the Elo score. Test that first. Make ten assets, not one, and lay them side by side.
What to do this week
Six steps. It takes about ninety minutes.
- Open grok.com/imagine, switch to Quality Mode, and run one real brief you already delivered last month.
- Test consistency, not beauty. Make one character or product, then place it in five settings. Line the results up.
- Run the background removal against your current cutout tool on ten messy images. Count the failures on each side.
- Try the multi-reference edit with a real composite. Five inputs, one output. Time it against your normal process.
- Try two Marketing templates that match a service you sell, such as E-Commerce Photos or Editorial Product Poster. Judge the output as a client would.
- Write down what you would never let it touch. Faces of real people, licensed assets, regulated claims. Decide that before a deadline decides it for you.
Keep the outputs. When the API arrives, that folder becomes your evidence for what to automate.
The limits to keep in mind
Be honest about the gaps.
It is second, not first. If you want the current leader on Arena, that is OpenAI's model.
The ranked variant is the low setting, and Arena rankings move week to week. Treat any leaderboard claim as a snapshot with a date on it.
There is no API and no published price, so you cannot forecast cost.
And the templates that read as marketing wins carry real risk. A "UGC Photos" template makes an image that looks like a customer took it. Disclosure rules and platform rules are yours to manage, not the model's. We covered the same trap when Google shipped three Gemini models in one day. New capability arrives faster than new policy.

The YARD take
Image models spent two years being judged on taste. This launch is being sold on labour.
Precise edits. Clean cutouts. Five-image composites. Nine ratios. Fifteen prebuilt workflows named after the work agencies bill for.
None of that makes a designer redundant. It moves the value. If your studio's advantage was resizing and retouching faster than the studio next door, that advantage is thinning. If your advantage is judgment and brand sense, it just got more valuable. Knowing which of the forty versions is the right one is still a human call.
So do not migrate anything on a launch-day claim. Do run the test this week. And do notice the direction of travel, because the next release will not be pointed at the picture either.
FAQ
What is Grok Imagine Image 2.0? It is SpaceXAI's next-generation image model, released on August 7, 2026. It runs as the new Quality Mode on grok.com/imagine and in the Grok iOS and Android apps.
Where does it rank against other image models? SpaceXAI says it is second in the world on both the Arena image-edit and text-to-image boards. OpenAI's GPT-Image-2 is first on both, with an Elo of 1,463 against 1,439 in editing and 1,380 against 1,320 in text-to-image.
Is there an API? Not yet. SpaceXAI says API access is coming soon. Today it is a consumer product only, so you cannot batch or automate it.
How many reference images can it take? Up to five in a single generation. That removes the manual compositing step for many jobs.
What is smart resize? You pick an aspect ratio and the model recomposes the frame to fit. Nine ratios ship at launch, from 9:16 through to 2:1.
Which templates matter for marketing teams? Four are filed under Marketing: Editorial Product Poster, E-Commerce Photos, UGC Photos and Merch Maker. Others, such as Professional Headshot and BG Removal and Change, are useful too.
What should we test first? Consistency. Generate one character or product, then place it across five scenes and compare. Brand work needs a repeatable look far more than it needs one great image.
Sources
- SpaceXAI — Imagine Image 2.0 (August 7, 2026; primary source for features, templates and Arena claim)
- The Decoder — xAI's Imagine Image 2.0 lands just behind OpenAI's GPT-Image-2 in Arena benchmarks (August 8, 2026; Elo figures)
- Unite.AI — xAI ships Grok Imagine Image 2.0 with precise editing and a top Arena ranking (August 2026)
- Product page — grok.com/imagine
Insights from Our Experts
Explore our latest articles on digital marketing strategies.




