Choose Auto, TokenLab Verified, or Official for each request, with prices shown up front.See what's new

xAI: Grok 4.3

xAI's Grok for chat, coding, agentic tools, and lower hallucination risk
Compare models
grok-4.3
AvailablexAIChatCached input 84% less
Input / Output-50%
$1.25 / $2.50$0.625 / $1.25
Context
1M
Max output
128K
Modalities
VisionChat
Capabilities
Tool usePrompt CacheReasoning

About Grok 4.3

Grok 4.3 is an xAI chat model that xAI describes as fast and reliable, with strong tool calling and instruction following. It reads text and images, writes text, and keeps the one-million-token class of context shared with the 4.20 family. Its reasoning effort is adjustable from none up to xhigh, so one model ID can cover quick replies and harder questions.

Where it works well

  • Reasoning effort ranges from none to xhigh, letting you trade speed for depth per request instead of switching models.
  • xAI positions it around reliable tool calling and instruction following, a fit for agents that must follow a fixed procedure.
  • Supports function calling and structured outputs for integrations that need machine-readable replies.
  • Takes images with text, so one request can combine a screenshot and a written question.
  • Prompt caching reduces repeat cost when a long system prompt or document is reused.

When to choose another model

  • Grok 4.5 through 4.7 are the newer releases and are aimed at coding and agentic software work; this one is the older, general-purpose pick.
  • At low effort it spends little time thinking, so hard problems need a higher setting or a stronger model.
  • Output is limited to text; image and audio generation need separate models.

Getting started

  1. Create API key

    Create a key in Console, then use it with every model on the platform.

  2. Send your first request

    Copy the example for your language and run it against the endpoint.

    Text to textResponses API
    POST/v1/responses
    API format:
    curl https://api.tokenlab.sh/v1/responses \
      -H "Content-Type: application/json" \
      -H "Authorization: Bearer sk-xxx" \
      -d '{
        "model": "grok-4.3",
        "input": "Hello!"
      }'

Pricing

Official price is the model maker's public baseline. TokenLab price is what you pay for this model on TokenLab.

Default spec

per 1M tokens
Official price
Input $1.25 / Output $2.50 / Cache read $0.20
TokenLab price
Input $0.625 / Output $1.25 / Cache read $0.10
Discount
-50%
Prompt cache pricing

Cache read

Official price
$0.20
TokenLab price
$0.10
Discount
-50%

Usage & activity

Success rate is the share of requests that completed. Latency is how long a full response takes; P95 means 95% of requests finished within that time.

Usage & availability

Last 24 hours
Requests
Success rate
P95 latency
Total tokens
Model performance
Metrics appear once privacy and data volume thresholds are met.

Data is based on aggregate user requests, excluding status checks.

Open in Console

Open Grok 4.3 in Console with a prompt ready to edit or send.

Help me try grok-4.3 with a short message at /v1/responses. Show the reply, latency, and cost.

Use cases

Best for
  • Reasoning
  • Vision
  • General-purpose assistant

    Back a chat feature with one model that answers simple questions fast and can be given more reasoning effort for tougher ones.

  • Instruction-bound agents

    Run workflows with a fixed procedure and several tools, such as triaging tickets, where the model has to follow steps exactly and call the right function.

  • Long-document question answering

    Load a large set of manuals or reports into the prompt and ask questions across them, with cached prefixes keeping repeat queries cheaper.

Prompt examples

Follow these five triage rules exactly. For each ticket below, call set_priority and assign_team, then give me a one-line reason.

Using the attached manuals, explain how to reset the controller after a fault, and quote the section you used.

Classify this review as bug, feature request or praise. Reply as JSON with a label and a confidence from 0 to 1.

This model has conditional pricing. Monthly token totals alone cannot produce a reliable estimate; use the detailed pricing for the request specification, cache, and applicable time window.

FAQ

What is Grok 4.3 designed for?

xAI describes it as a fast, reliable model with strong tool calling and instruction following. It suits assistants and agents that must follow a procedure, call functions, and return structured answers, over text and image input.

Does Grok 4.3 support reasoning?

Yes, with an adjustable effort level of none, low, medium, high or xhigh. xAI's default is low. Raise the level for harder questions and set it to none when you want the quickest reply.

How does Grok 4.3 compare with Grok 4.20?

Both share a very long context and similar input types. Grok 4.3 adds a finer reasoning-effort control in one ID, whereas 4.20 is split into separate reasoning, non-reasoning and multi-agent models. Test both on your own prompts.

Can Grok 4.3 process images?

Yes. It accepts images alongside text, such as charts and screenshots, and answers in text. It does not generate images. That makes it handy for questions about a screenshot or a photographed page, and the answer arrives as ordinary text.

Is Grok 4.3 good for coding?

It can handle everyday code questions and tool-driven edits. For demanding software engineering and long agent runs, xAI's newer Grok 4.5, 4.6 and 4.7 are positioned more directly at coding and agentic tasks.

How much does Grok 4.3 cost?

On TokenLab, Grok 4.3 costs Input $0.625 / Output $1.25 per 1M tokens. The pricing table above shows the full breakdown. Rates depend on the billing unit, specification, and usage. Compare matching conditions in the model's detailed pricing; a single rate does not determine the total cost.

What are the context window and output limit of Grok 4.3?

Grok 4.3 accepts up to 1,000,000 tokens of context and returns up to 131,072 tokens in one response.

Which endpoint should Grok 4.3 use?

Use https://api.tokenlab.sh/v1/responses for Grok 4.3. The request example below shows the matching code shape.

Which operations does Grok 4.3 support?

Grok 4.3 supports Text to text. Select an operation above to see its endpoint and request example.

Compare Grok 4.3

Sources

Reviewed Oct 2, 2026

More from Grok 4

Related models