xAI: Grok 4.3
grok-4.3- Input / Output-50%
- $1.25 / $2.50$0.625 / $1.25
- Context
- 1M
- Max output
- 128K
- Modalities
- VisionChat
- Capabilities
- Tool usePrompt CacheReasoning
About Grok 4.3
Grok 4.3 is an xAI chat model that xAI describes as fast and reliable, with strong tool calling and instruction following. It reads text and images, writes text, and keeps the one-million-token class of context shared with the 4.20 family. Its reasoning effort is adjustable from none up to xhigh, so one model ID can cover quick replies and harder questions.
Where it works well
- Reasoning effort ranges from none to xhigh, letting you trade speed for depth per request instead of switching models.
- xAI positions it around reliable tool calling and instruction following, a fit for agents that must follow a fixed procedure.
- Supports function calling and structured outputs for integrations that need machine-readable replies.
- Takes images with text, so one request can combine a screenshot and a written question.
- Prompt caching reduces repeat cost when a long system prompt or document is reused.
When to choose another model
- Grok 4.5 through 4.7 are the newer releases and are aimed at coding and agentic software work; this one is the older, general-purpose pick.
- At low effort it spends little time thinking, so hard problems need a higher setting or a stronger model.
- Output is limited to text; image and audio generation need separate models.
Getting started
Create API key
Create a key in Console, then use it with every model on the platform.
Send your first request
Copy the example for your language and run it against the endpoint.
Text to textResponses APIPOST/v1/responsesAPI format:curl https://api.tokenlab.sh/v1/responses \ -H "Content-Type: application/json" \ -H "Authorization: Bearer sk-xxx" \ -d '{ "model": "grok-4.3", "input": "Hello!" }'
Pricing
Official price is the model maker's public baseline. TokenLab price is what you pay for this model on TokenLab.
Default spec
per 1M tokens- Official price
- Input $1.25 / Output $2.50 / Cache read $0.20
- TokenLab price
- Input $0.625 / Output $1.25 / Cache read $0.10
- Discount
- -50%
Cache read
- Official price
- $0.20
- TokenLab price
- $0.10
- Discount
- -50%
| Official priceper 1M tokens | TokenLab priceper 1M tokens | Discount | |
|---|---|---|---|
| Default spec | Input $1.25 / Output $2.50 / Cache read $0.20 | Input $0.625 / Output $1.25 / Cache read $0.10 | -50% |
| Prompt cache pricing | |||
| Cache read | $0.20 | $0.10 | -50% |
Usage & activity
Success rate is the share of requests that completed. Latency is how long a full response takes; P95 means 95% of requests finished within that time.
Usage & availability
Last 24 hours- Requests
- Success rate
- P95 latency
- Total tokens
Data is based on aggregate user requests, excluding status checks.
Open in Console
Open Grok 4.3 in Console with a prompt ready to edit or send.
Help me try grok-4.3 with a short message at /v1/responses. Show the reply, latency, and cost.
Use cases
Best for- Reasoning
- Vision
General-purpose assistant
Back a chat feature with one model that answers simple questions fast and can be given more reasoning effort for tougher ones.
Instruction-bound agents
Run workflows with a fixed procedure and several tools, such as triaging tickets, where the model has to follow steps exactly and call the right function.
Long-document question answering
Load a large set of manuals or reports into the prompt and ask questions across them, with cached prefixes keeping repeat queries cheaper.
Prompt examples
Follow these five triage rules exactly. For each ticket below, call set_priority and assign_team, then give me a one-line reason.
Using the attached manuals, explain how to reset the controller after a fault, and quote the section you used.
Classify this review as bug, feature request or praise. Reply as JSON with a label and a confidence from 0 to 1.
This model has conditional pricing. Monthly token totals alone cannot produce a reliable estimate; use the detailed pricing for the request specification, cache, and applicable time window.
FAQ
What is Grok 4.3 designed for?
xAI describes it as a fast, reliable model with strong tool calling and instruction following. It suits assistants and agents that must follow a procedure, call functions, and return structured answers, over text and image input.
Does Grok 4.3 support reasoning?
Yes, with an adjustable effort level of none, low, medium, high or xhigh. xAI's default is low. Raise the level for harder questions and set it to none when you want the quickest reply.
How does Grok 4.3 compare with Grok 4.20?
Both share a very long context and similar input types. Grok 4.3 adds a finer reasoning-effort control in one ID, whereas 4.20 is split into separate reasoning, non-reasoning and multi-agent models. Test both on your own prompts.
Can Grok 4.3 process images?
Yes. It accepts images alongside text, such as charts and screenshots, and answers in text. It does not generate images. That makes it handy for questions about a screenshot or a photographed page, and the answer arrives as ordinary text.
Is Grok 4.3 good for coding?
It can handle everyday code questions and tool-driven edits. For demanding software engineering and long agent runs, xAI's newer Grok 4.5, 4.6 and 4.7 are positioned more directly at coding and agentic tasks.
How much does Grok 4.3 cost?
On TokenLab, Grok 4.3 costs Input $0.625 / Output $1.25 per 1M tokens. The pricing table above shows the full breakdown. Rates depend on the billing unit, specification, and usage. Compare matching conditions in the model's detailed pricing; a single rate does not determine the total cost.
What are the context window and output limit of Grok 4.3?
Grok 4.3 accepts up to 1,000,000 tokens of context and returns up to 131,072 tokens in one response.
Which endpoint should Grok 4.3 use?
Use https://api.tokenlab.sh/v1/responses for Grok 4.3. The request example below shows the matching code shape.
Which operations does Grok 4.3 support?
Grok 4.3 supports Text to text. Select an operation above to see its endpoint and request example.
Compare Grok 4.3
Sources
Reviewed Oct 2, 2026