Choose Auto, TokenLab Verified, or Official for each request, with prices shown up front.See what's new

xAI: Grok 4.7

Grok 4.7 focuses on coding, agentic tasks, and knowledge work. Reasoning, image understanding, and tool calling make it suitable for assistants that need to diagnose a problem, act on it, and evaluate the result.
Compare models
grok-4.7
AvailablexAIChatCached input 75% lessNew
Input / Output-50%
$2.00 / $6.00$1.00 / $3.00
Context
500K
Released
Sep 21, 2026
Max output
128K
Modalities
VisionChat
Capabilities
Tool usePrompt CacheReasoning

About Grok 4.7

Grok 4.7 is xAI's frontier model for coding, agentic tasks and knowledge work, and the one xAI's documentation recommends first. It reads text and images, writes text, calls tools and returns structured output, with reasoning effort from low to xhigh. It suits assistants that diagnose a problem, act on it and then evaluate the result.

Where it works well

  • xAI recommends it as its most capable model, aimed at coding and agentic loops that diagnose, act and confirm the outcome.
  • Reasoning effort runs from low to xhigh, with high as the default for demanding work.
  • Accepts images with text, so a visual bug or a chart can be part of the task.
  • Function calling and structured outputs support multi-tool agents and strict data formats.
  • Reasoning is always performed, which benefits tasks that need planning before action.

When to choose another model

  • Always-on reasoning makes it slower than the non-reasoning 4.20 model for simple drafting or extraction.
  • It cannot generate images, audio or video; output is text.
  • Its context window is smaller than that of Grok 4.3 and the 4.20 family, which matters for very large inputs.

Getting started

  1. Create API key

    Create a key in Console, then use it with every model on the platform.

  2. Send your first request

    Copy the example for your language and run it against the endpoint.

    Text to textResponses API
    POST/v1/responses
    API format:
    curl https://api.tokenlab.sh/v1/responses \
      -H "Content-Type: application/json" \
      -H "Authorization: Bearer sk-xxx" \
      -d '{
        "model": "grok-4.7",
        "input": "Hello!"
      }'

Pricing

Official price is the model maker's public baseline. TokenLab price is what you pay for this model on TokenLab.

Input tokens <= 200K

per 1M tokens
Official price
Input $2.00 / Output $6.00 / Cache read $0.50
TokenLab price
Input $1.00 / Output $3.00 / Cache read $0.25
Discount
-50%

Input tokens 200K-500K

per 1M tokens
Official price
Input $4.00 / Output $12.00 / Cache read $1.00
TokenLab price
Input $2.00 / Output $6.00 / Cache read $0.50
Discount
-50%
Prompt cache pricing

Cache read

Official price
$0.50
TokenLab price
$0.25
Discount
-50%

Usage & activity

Success rate is the share of requests that completed. Latency is how long a full response takes; P95 means 95% of requests finished within that time.

Usage & availability

Last 24 hours
Requests
Success rate
P95 latency
Total tokens
Model performance
Metrics appear once privacy and data volume thresholds are met.

Data is based on aggregate user requests, excluding status checks.

Open in Console

Open Grok 4.7 in Console with a prompt ready to edit or send.

Help me try grok-4.7 with a short message at /v1/responses. Show the reply, latency, and cost.

Use cases

Best for
  • Reasoning
  • Vision
  • Autonomous coding agents

    Let it read a repository, make multi-file edits, run tests through a tool, interpret failures and iterate until the change passes.

  • Incident diagnosis

    Give it alerts, logs and dashboard screenshots, and have it find the cause, apply a documented fix through tools and confirm recovery.

  • Complex knowledge work

    Prepare analyses that combine long source material, calculations and a written recommendation, with the reasoning level raised for the hardest questions.

Prompt examples

The checkout test fails on CI only. Inspect the repo, find why, patch it, run the test, and report what you changed.

These three dashboards and the error log cover one outage. Identify the first failing component and propose a rollback plan.

Review this pricing model spreadsheet export and the memo. List where the memo's conclusions do not follow from the numbers.

This model has conditional pricing. Monthly token totals alone cannot produce a reliable estimate; use the detailed pricing for the request specification, cache, and applicable time window.

FAQ

What is Grok 4.7 good at?

xAI describes it as built for coding, agentic tasks and knowledge work. It combines reasoning, image understanding and tool calling, so it can diagnose a problem, take action through tools, and evaluate the outcome in one loop.

What reasoning levels does Grok 4.7 have?

Low, medium, high and xhigh, with high as the default. Reasoning is always on for this model, so the setting controls how much effort it spends rather than whether it thinks.

Can Grok 4.7 read screenshots and charts?

Yes. It accepts images with text, which helps with UI bugs, dashboards and figures in documents. Output remains text. Ask a text question such as why a layout breaks or what a graph shows, and the answer comes back as written explanation.

Is Grok 4.7 better than Grok 4.6?

Grok 4.7 is the later release and xAI's recommended model, with the same inputs and effort levels. Expect it to be the stronger starting point for agentic work.

Does Grok 4.7 support function calling?

Yes, along with structured outputs. You can define tools and JSON schemas in the request, which makes it suitable for agents that need to act on external systems.

How much does Grok 4.7 cost?

On TokenLab, Grok 4.7 costs Input $1.00 / Output $3.00 per 1M tokens. The pricing table above shows the full breakdown. Rates depend on the billing unit, specification, and usage. Compare matching conditions in the model's detailed pricing; a single rate does not determine the total cost.

What are the context window and output limit of Grok 4.7?

Grok 4.7 accepts up to 500,000 tokens of context and returns up to 131,072 tokens in one response.

Which endpoint should Grok 4.7 use?

Use https://api.tokenlab.sh/v1/responses for Grok 4.7. The request example below shows the matching code shape.

Which operations does Grok 4.7 support?

Grok 4.7 supports Text to text. Select an operation above to see its endpoint and request example.

Compare Grok 4.7

Sources

Reviewed Oct 2, 2026

More from Grok 4

Related models