Choose Auto, TokenLab Verified, or Official for each request, with prices shown up front.See what's new

xAI: Grok Imagine Image

Grok Imagine Image covers both new image creation and edits to supplied pictures. Its range of aspect ratios is useful when preparing the same creative concept for social posts, banners, and portrait-oriented designs.
Compare models
grok-imagine-image
AvailablexAIImageSync
Price
From $0.02
Modalities
Image

About Grok Imagine Image

Grok Imagine Image is xAI's image model for generating pictures from text and for editing pictures you supply. It is the standard-tier Imagine model, placed below the higher-fidelity quality model and the newer 2.0 release. It offers 1K and 2K resolution and a wide set of aspect ratios, including tall phone formats, which helps when one idea must fit posts, banners and stories.

Where it works well

  • Covers both text-to-image generation and editing of an existing image in the same model.
  • Offers many aspect ratios, from square to wide and tall phone shapes, plus an auto option.
  • Choice of 1K or 2K output lets you draft small and render larger once a concept works.
  • Standard tier suits drafts and volume work where the quality variant is not needed.
  • Natural-language edit instructions change a supplied picture without re-describing the whole scene.

When to choose another model

  • The quality and 2.0 models are the better picks when fine detail matters.
  • It produces still images only; moving shots need a Grok Imagine video model.
  • Text rendered inside images and exact layouts can need several attempts.

Getting started

  1. Create API key

    Create a key in Console, then use it with every model on the platform.

  2. Send your first request

    Copy the example for your language and run it against the endpoint.

    Image editingImage to imageText to imageTokenLab endpoint
    POST/v1/images/generations
    curl -X POST "https://api.tokenlab.sh/v1/images/generations" \
      -H "Authorization: Bearer sk-xxx" \
      -H "Content-Type: application/json" \
      -d '{
      "operation": "text-to-image",
      "model": "grok-imagine-image",
      "prompt": "A minimalist product photo of matte black headphones on a soft blue background."
    }'

Pricing

Official price is the model maker's public baseline. TokenLab price is what you pay for this model on TokenLab.

Pricing

per request
TokenLab price
How pricing works
Official price
How pricing works

Usage & activity

Success rate is the share of requests that completed. Latency is how long a full response takes; P95 means 95% of requests finished within that time.

Usage & availability

Last 24 hours
Requests
Success rate
P95 latency
Total tokens
Model performance
Metrics appear once privacy and data volume thresholds are met.

Data is based on aggregate user requests, excluding status checks.

Open in Console

Open Grok Imagine Image in Console with a prompt ready to edit or send.

Help me create an image with grok-imagine-image using text-to-image at /v1/images/generations. Show the result, cost, and any limits I should know.

Use cases

Best for
  • Text to Image
  • Social media variants

    Render one concept at square, portrait and widescreen ratios so the same visual fits a feed post, a story and a banner.

  • Concept drafts

    Explore many visual directions quickly from short prompts, then move the best one to a higher-fidelity model for final art.

  • Photo edits by instruction

    Upload a product photo and ask for a new background or a changed colour, keeping the subject in place.

Prompt examples

A vertical poster of a lighthouse at dusk, painted in thick oil strokes, with space at the top for a title.

Change the background of this product photo to a pale concrete studio wall and keep the shadows natural.

Three-by-two banner for a coffee brand: steam rising from a cup on a wooden table, warm morning light.

FAQ

What can Grok Imagine Image do?

It generates images from a text prompt and edits images you provide, following a written instruction. You can choose a 1K or 2K output size and pick from many aspect ratios, including portrait and widescreen.

How does grok-imagine-image differ from the quality model?

This is xAI's standard Imagine image tier, and grok-imagine-image-quality is the higher-fidelity one. Start here for drafts and larger batches, and move to the quality model when you need finer detail in the final picture.

Which aspect ratios does it support?

It supports square, 3:2, 4:3, 16:9 and their portrait counterparts, plus taller phone shapes such as 9:19.5 and 9:20, and an auto setting that lets the model choose.

Can it edit my own images?

Yes. Supply an image and a text instruction, and the model returns an edited version. This is useful for background swaps, colour changes and small additions to an existing picture.

Does it generate video?

No, it makes still images. For motion, use one of the Grok Imagine video models, which can animate an image or generate a clip from text.

How much does Grok Imagine Image cost?

On TokenLab, Grok Imagine Image costs How pricing works . The pricing table above shows the full breakdown. Rates depend on the billing unit, specification, and usage. Compare matching conditions in the model's detailed pricing; a single rate does not determine the total cost.

Which endpoint should Grok Imagine Image use?

Use https://api.tokenlab.sh/v1/images/generations for Grok Imagine Image. The request example below shows the matching code shape.

Which operations does Grok Imagine Image support?

Grok Imagine Image supports Image editing, Image to image, Text to image. Select an operation above to see its endpoint and request example.

Compare Grok Imagine Image

Guides that use Grok Imagine Image

Sources

Reviewed Oct 2, 2026

More from Grok Imagine

Related models