Grok 4.6
import { streamText } from 'ai'
const result = streamText({ model: 'spacexai/grok-4.6', prompt: 'Why is the sky blue?'})Copy link to headingPlayground
Try out Grok 4.6 by SpaceXAI. Usage is billed to your team at API rates. Free users (those who haven't made a payment) get $5 of credits every 30 days.
Grok 4.6
Copy link to headingProviders
Route requests across multiple providers. Copy a provider slug to set your preference. Visit the docs for more info. Using a provider means you agree to their terms, listed under Legal.
| Provider |
|---|
Copy link to headingThroughput24 hours
P50 throughput on live AI Gateway traffic, in tokens per second (TPS). Visit the docs for more info.
Copy link to headingLatency24 hours
P50 time to first token (TTFT) on live AI Gateway traffic, in milliseconds. View the docs for more info.
Copy link to headingUptime24 hours
Direct request success rate on AI Gateway and per-provider. Visit the docs for more info.
Copy link to headingAbout Grok 4.6
Grok 4.6 was released August 12, 2026 as SpaceXAI's frontier reasoning model for long-running agents, coding, and knowledge work. It refines Grok 4.5 rather than replacing the foundation underneath it, so the shape of the model is familiar and the gains show up in sustained agentic work. Grok 4.6 accepts text and image input, returns text, and works within a context window of 500K tokens. SpaceXAI lists a knowledge cutoff of February 1, 2026.
Reasoning depth is a per-request setting, and Grok 4.6 adds a fourth level. Alongside low, medium, and high, an xhigh setting pushes the model further on problems that reward more deliberation. High remains the default. On longer trajectories Grok 4.6 checks its own work before moving on, which is the behaviour that matters most when an agent chains dozens of steps without a human in the loop.
Grok 4.6 leads on GDPval-AA v2 and AA-Briefcase, and scores 61 on the Artificial Analysis Intelligence Index. Coding results are more mixed: it trails the strongest models on DeepSWE v1.1 and Terminal-Bench v3.0, so benchmark it against your own repository before switching a coding agent over.
Pricing is tiered by prompt length, with a higher rate once a request crosses 200K tokens. See the pricing panel on this page for current rates before you send very long prompts.
You can integrate Grok 4.6 through AI SDK, Chat Completions API, Responses API, Messages API, or other API formats, from TypeScript or Python. Routing rules move traffic from another Grok model to Grok 4.6 without changing application code.
Copy link to headingWhat To Consider When Choosing a Provider
- Configuration: Reasoning stays on for every Grok 4.6 request, so responses carry thinking tokens even at the low level. The new
xhighlevel spends more of them than high, which is the point, but it makes per-request output cost harder to predict. Set the level per request rather than leaving everything at the default. - Configuration: Prompt length changes the rate. Requests at or above 200K tokens bill at a higher tier, and that tier applies to the whole request rather than only the tokens past the threshold, so a prompt that drifts just over the line costs more than one that stays just under it.
- Configuration: Coding is the weaker column on the launch benchmarks. If your workload is mostly software engineering rather than general knowledge work, compare Grok 4.6 against the alternatives below on your own tasks. The February 1, 2026 knowledge cutoff also means Grok 4.6 needs web search or retrieval for anything more recent.
- Zero Data Retention: Zero Data Retention is offered on a per-provider and model basis. See the documentation for details.
- Authentication: AI Gateway authenticates requests using an API key or OIDC token. You do not need to manage provider credentials directly.
Copy link to headingWhen to Use Grok 4.6
Best for
- Long-Running Agents: Multi-hour trajectories where the model verifies its own work between steps
- Extended Reasoning Depth: A fourth
xhighlevel for problems that reward more deliberation - Knowledge Work: Research, analysis, and briefing tasks where it leads on GDPval-AA v2
- Large Context Tasks: A 500K tokens window that holds long documents and full session history
- Vision-Assisted Work: Screenshots, diagrams, and design references alongside text prompts
Consider alternatives when
- Heavy Coding Workloads: Grok Code Fast 1 is tuned for quick code edits at lower cost
- Reasoning-Free Responses: Grok 4.1 Fast Non-Reasoning answers without spending thinking tokens
- Short Prompt Volume: Grok 4.5 avoids the long-prompt pricing tier on smaller requests
- Multi-Agent Orchestration: Grok 4.20 Multi-Agent spreads one request across collaborating agents
Copy link to headingConclusion
Grok 4.6 is SpaceXAI's frontier model for agents that run long, with a 500K tokens window, four reasoning levels, and self-verification across extended trajectories. Point spacexai/grok-4.6 at AI Gateway to route requests behind one API key, and watch the 200K prompt threshold where the higher pricing tier begins.