xAI: Grok Imagine Image
grok-imagine-image- Price
- From $0.02
- Modalities
- Image
About Grok Imagine Image
Grok Imagine Image is xAI's image model for generating pictures from text and for editing pictures you supply. It is the standard-tier Imagine model, placed below the higher-fidelity quality model and the newer 2.0 release. It offers 1K and 2K resolution and a wide set of aspect ratios, including tall phone formats, which helps when one idea must fit posts, banners and stories.
Where it works well
- Covers both text-to-image generation and editing of an existing image in the same model.
- Offers many aspect ratios, from square to wide and tall phone shapes, plus an auto option.
- Choice of 1K or 2K output lets you draft small and render larger once a concept works.
- Standard tier suits drafts and volume work where the quality variant is not needed.
- Natural-language edit instructions change a supplied picture without re-describing the whole scene.
When to choose another model
- The quality and 2.0 models are the better picks when fine detail matters.
- It produces still images only; moving shots need a Grok Imagine video model.
- Text rendered inside images and exact layouts can need several attempts.
Getting started
Create API key
Create a key in Console, then use it with every model on the platform.
Send your first request
Copy the example for your language and run it against the endpoint.
Image editingImage to imageText to imageTokenLab endpointPOST/v1/images/generationscurl -X POST "https://api.tokenlab.sh/v1/images/generations" \ -H "Authorization: Bearer sk-xxx" \ -H "Content-Type: application/json" \ -d '{ "operation": "text-to-image", "model": "grok-imagine-image", "prompt": "A minimalist product photo of matte black headphones on a soft blue background." }'
Pricing
Official price is the model maker's public baseline. TokenLab price is what you pay for this model on TokenLab.
Pricing
per request- TokenLab price
- How pricing works
- Official price
- How pricing works
| Official priceper request | TokenLab priceper request | Discount | |
|---|---|---|---|
| Price | How pricing works | How pricing works | — |
Usage & activity
Success rate is the share of requests that completed. Latency is how long a full response takes; P95 means 95% of requests finished within that time.
Usage & availability
Last 24 hours- Requests
- Success rate
- P95 latency
- Total tokens
Data is based on aggregate user requests, excluding status checks.
Open in Console
Open Grok Imagine Image in Console with a prompt ready to edit or send.
Help me create an image with grok-imagine-image using text-to-image at /v1/images/generations. Show the result, cost, and any limits I should know.
Use cases
Best for- Text to Image
Social media variants
Render one concept at square, portrait and widescreen ratios so the same visual fits a feed post, a story and a banner.
Concept drafts
Explore many visual directions quickly from short prompts, then move the best one to a higher-fidelity model for final art.
Photo edits by instruction
Upload a product photo and ask for a new background or a changed colour, keeping the subject in place.
Prompt examples
A vertical poster of a lighthouse at dusk, painted in thick oil strokes, with space at the top for a title.
Change the background of this product photo to a pale concrete studio wall and keep the shadows natural.
Three-by-two banner for a coffee brand: steam rising from a cup on a wooden table, warm morning light.
FAQ
What can Grok Imagine Image do?
It generates images from a text prompt and edits images you provide, following a written instruction. You can choose a 1K or 2K output size and pick from many aspect ratios, including portrait and widescreen.
How does grok-imagine-image differ from the quality model?
This is xAI's standard Imagine image tier, and grok-imagine-image-quality is the higher-fidelity one. Start here for drafts and larger batches, and move to the quality model when you need finer detail in the final picture.
Which aspect ratios does it support?
It supports square, 3:2, 4:3, 16:9 and their portrait counterparts, plus taller phone shapes such as 9:19.5 and 9:20, and an auto setting that lets the model choose.
Can it edit my own images?
Yes. Supply an image and a text instruction, and the model returns an edited version. This is useful for background swaps, colour changes and small additions to an existing picture.
Does it generate video?
No, it makes still images. For motion, use one of the Grok Imagine video models, which can animate an image or generate a clip from text.
How much does Grok Imagine Image cost?
On TokenLab, Grok Imagine Image costs How pricing works . The pricing table above shows the full breakdown. Rates depend on the billing unit, specification, and usage. Compare matching conditions in the model's detailed pricing; a single rate does not determine the total cost.
Which endpoint should Grok Imagine Image use?
Use https://api.tokenlab.sh/v1/images/generations for Grok Imagine Image. The request example below shows the matching code shape.
Which operations does Grok Imagine Image support?
Grok Imagine Image supports Image editing, Image to image, Text to image. Select an operation above to see its endpoint and request example.
Compare Grok Imagine Image
Guides that use Grok Imagine Image
Sources
Reviewed Oct 2, 2026