Chat Models
AI models for chat, Q&A, reasoning, and text generation
Chat · xAI
CodeJSON modePrompt CacheReasoningTool useVisionCached inputless
- TokenLab price— Discount
- —
- Official price
- Input$1.00Output$2.00/1M
JSON modePrompt CacheReasoningTool useVisionCached input 84% less
- TokenLab price-50% Discount
- Input$0.625Output$1.25/1M
- Official price
- Input$1.25Output$2.50/1M
JSON modePrompt CacheReasoningTool useVisionCached inputless
- TokenLab price— Discount
- —
- Official price
- Input$1.25Output$2.50/1M
JSON modePrompt CacheTool useVisionCached input 84% less
- TokenLab price-50% Discount
- Input$0.625Output$1.25/1M
- Official price
- Input$1.25Output$2.50/1M
JSON modePrompt CacheReasoningTool useVisionCached input 84% less
- TokenLab price-50% Discount
- Input$0.625Output$1.25/1M
- Official price
- Input$1.25Output$2.50/1M
JSON modePrompt CacheReasoningTool useVisionCached input 75% less
- TokenLab price-50% Discount
- Input$1.00Output$3.00/1M
- Official price
- Input$2.00Output$6.00/1M
JSON modePrompt CacheTool useVisionCached input 75% less
- TokenLab price-50% Discount
- Input$1.00Output$3.00/1M
- Official price
- Input$2.00Output$6.00/1M
All model series
1 seriesChoosing the right chat model
The right model balances task fit, output quality, latency, and price.
Selection signals
- Match the model to the input and output you actually need.
- Compare models that use the same pricing unit.
- A small test reveals quality and latency differences that a catalog cannot.
FAQ
What makes a good chat model?
A strong fit matches the task, quality bar, latency target, and budget. A small evaluation set makes quality and reliability differences easier to see.
Can I switch models later?
Yes. Keep the public model ID and request format explicit, and compare alternatives on the same tasks.