xAI: Grok 4.5
grok-4.5- Input / Output-50%
- $2.00 / $6.00$1.00 / $3.00
- Context
- 500K
- Released
- Jul 8, 2026
- Max output
- 128K
- Modalities
- VisionChat
- Capabilities
- Tool usePrompt CacheReasoning
About Grok 4.5
Grok 4.5 is xAI's model for agentic software, engineering and workflow tasks, trained with an emphasis on science, math and engineering material. It accepts text and images, writes text, and reasons at four adjustable effort levels with high as the default. It sits between Grok 4.3 and the later 4.6 and 4.7 releases, with a smaller context window than 4.3.
Where it works well
- Training emphasis on science, math and engineering data makes it a sensible pick for technical reasoning questions.
- Four reasoning effort levels, low to xhigh, let you tune depth against latency on each request.
- Function calling and structured outputs support agents that edit code or drive internal tools.
- Reads screenshots and diagrams along with text, which helps with UI bugs and technical figures.
- Prompt caching suits coding sessions that resend the same repository context.
When to choose another model
- Its context window is smaller than that of Grok 4.3 and the 4.20 family, so very large corpora may not fit.
- Grok 4.6 and 4.7 are later releases; 4.7 is xAI's current recommended model for coding.
- Text-only output means no generated images, audio or video.
Getting started
Create API key
Create a key in Console, then use it with every model on the platform.
Send your first request
Copy the example for your language and run it against the endpoint.
Text to textResponses APIPOST/v1/responsesAPI format:curl https://api.tokenlab.sh/v1/responses \ -H "Content-Type: application/json" \ -H "Authorization: Bearer sk-xxx" \ -d '{ "model": "grok-4.5", "input": "Hello!" }'
Pricing
Official price is the model maker's public baseline. TokenLab price is what you pay for this model on TokenLab.
Default spec
per 1M tokens- Official price
- Input $2.00 / Output $6.00 / Cache read $0.50
- TokenLab price
- Input $1.00 / Output $3.00 / Cache read $0.25
- Discount
- -50%
Cache read
- Official price
- $0.50
- TokenLab price
- $0.25
- Discount
- -50%
| Official priceper 1M tokens | TokenLab priceper 1M tokens | Discount | |
|---|---|---|---|
| Default spec | Input $2.00 / Output $6.00 / Cache read $0.50 | Input $1.00 / Output $3.00 / Cache read $0.25 | -50% |
| Prompt cache pricing | |||
| Cache read | $0.50 | $0.25 | -50% |
Usage & activity
Success rate is the share of requests that completed. Latency is how long a full response takes; P95 means 95% of requests finished within that time.
Usage & availability
Last 24 hours- Requests
- Success rate
- P95 latency
- Total tokens
Data is based on aggregate user requests, excluding status checks.
Open in Console
Open Grok 4.5 in Console with a prompt ready to edit or send.
Help me try grok-4.5 with a short message at /v1/responses. Show the reply, latency, and cost.
Use cases
Best for- Reasoning
- Vision
Coding agents
Plan a change, edit files through tool calls, run checks and revise, with reasoning effort raised for tricky refactors and lowered for routine fixes.
Technical problem solving
Work through physics, math or engineering questions where intermediate steps matter and the answer should be checkable.
Workflow automation
Chain tool calls across internal systems, such as reading a ticket, querying a database and drafting a response, with structured outputs between steps.
Prompt examples
Read the stack trace and the two files below. Find the root cause, patch it, and tell me which test would have caught it.
Derive the closed form for the time a damped oscillator takes to fall below 5 percent amplitude, and show each step.
Given this ticket, call lookup_account, then draft a reply. If the data conflicts with what the customer says, say so before replying.
This model has conditional pricing. Monthly token totals alone cannot produce a reliable estimate; use the detailed pricing for the request specification, cache, and applicable time window.
FAQ
What is Grok 4.5 best at?
xAI positions it for agentic software, engineering and workflow tasks, and says it was trained with an emphasis on science, math and engineering data. It fits coding agents and technical reasoning more than casual chat.
What reasoning levels does Grok 4.5 offer?
There are four: low, medium, high and xhigh, with high as the default. Higher levels give the model more room to think before answering, at the cost of longer replies and more tokens.
Does Grok 4.5 accept image input?
Yes. It takes text and images as input and returns text, so you can include a screenshot of an error or a diagram with your question.
Should I use Grok 4.5 or Grok 4.7?
Grok 4.7 is the later release and the one xAI recommends as its most capable. Choose 4.5 if your prompts are already tuned to it and quality is sufficient, and compare both on a sample of your own tasks.
Does it support function calling?
Yes, along with structured outputs. You define functions or a JSON schema in the request and the model can call the functions or return data in that shape.
How much does Grok 4.5 cost?
On TokenLab, Grok 4.5 costs Input $1.00 / Output $3.00 per 1M tokens. The pricing table above shows the full breakdown. Rates depend on the billing unit, specification, and usage. Compare matching conditions in the model's detailed pricing; a single rate does not determine the total cost.
What are the context window and output limit of Grok 4.5?
Grok 4.5 accepts up to 500,000 tokens of context and returns up to 131,072 tokens in one response.
Which endpoint should Grok 4.5 use?
Use https://api.tokenlab.sh/v1/responses for Grok 4.5. The request example below shows the matching code shape.
Which operations does Grok 4.5 support?
Grok 4.5 supports Text to text. Select an operation above to see its endpoint and request example.
Compare Grok 4.5
Sources
Reviewed Oct 2, 2026