Grok 4.20
Grok 4.20 is worth evaluating for coding, rag, and agents when its provider route and context window match the workload.
Use it for
- Teams evaluating coding, rag, and agents
- Workloads that can use a 1m context window
- Buyers comparing 2 tracked provider routes
Do not use it for
- Workloads where another current model has stronger sourced task evidence
- Family
- Grok 4
- Released
- 2026-02-17
- Context
- 1m
- Knowledge cutoff
- 2024-11
- Specialization
- general
- Openness
- Proprietary
- License
- ProprietaryCommercial use: conditional
- Weights
- Not released
- Code
- Unknown
Cheapest of 2 routes · OpenRouter
About
Grok 4.20 is xAI's February 2026 Grok 4-series model, first previewed under the informal Grok 4.2 beta label. Standard API variants launched around March 10, 2026 as grok-4.20-0309-reasoning and grok-4.20-0309-non-reasoning with a 1M context window.
Grok 4.20 is a proprietary model in the Grok 4 family. The structured metadata tracks a 1m-token context window, multimodal input, reasoning, function calling, tool use, and structured outputs. This page tracks provider routes through xAI Console and OpenRouter, with the cheapest tracked route listed at $1.25 input and $2.5 output per 1M tokens. Headline tracked benchmarks include τ-bench 78.9, SWE-bench Verified 76.7, and Google-Proof Q&A 88.0.
Top use-case fit: coding, agents, and build tasks
Coding
Q/$ C2 relevant benchmarks in the decision map.
RAG
Included by capability and metadata signals in the decision map.
Agents
Q/$ C2 relevant benchmarks in the decision map.
Provider price ladder
Compare all 2Compare API pricing across 2 providers for input and output tokens, batch, and cached reads when available.
| Provider | Input / 1M | Output / 1M | Route |
|---|---|---|---|
| OpenRouter | $1.25 | $2.50 | Serverless |
| xAI Console | $1.25 | $2.50 | Serverless |
Available via routers & gateways(1)
Capabilities
Benchmark peer barsfor Coding
Benchmark scores(6)
| Benchmark | Score | Version | Source |
|---|---|---|---|
| τ-bench | 78.9 | τ-bench | https://benchlm.ai/benchmarks/tauBench |
| SWE-bench Verified | 76.7 | SWE-bench Verified | https://benchlm.ai/benchmarks/sweVerified |
| Google-Proof Q&A | 88.0 | — | https://x.ai/blog/grok-4 |
| Aider Polyglot | 79.6 | Listed as 'Grok-4 (high)' = 79 (percent_correct) | https://aider.chat/docs/leaderboards/ |
| ARC-AGI-2 | 53.3 | ARC-AGI-2 (accuracy%) | https://benchlm.ai/benchmarks/arcAgi2 |
| Humanity's Last Exam | 32.2 | HLE (accuracy) | https://designforonline.com/ai-models/xai-grok-4-20/ |
Migration checks
No linked migration route is available for this model yet.
Rankings & picks(2)
Compare Grok 4.20 with other models
Comparison and alternatives
Browse all comparisons →Frequently asked questions
What is the context window of Grok 4.20?
Grok 4.20 has a context window of 1m tokens.
How much does Grok 4.20 cost?
Grok 4.20 is available at $1.25/1M input tokens through xAI Console.
When was Grok 4.20 released?
Grok 4.20 was released on 2026-02-17.
Which providers offer Grok 4.20?
Grok 4.20 is available from 2 providers: xAI Console, OpenRouter.
What benchmarks has Grok 4.20 been tested on?
Grok 4.20 has been evaluated on 6 benchmarks, including τ-bench, SWE-bench Verified, Google-Proof Q&A, Aider Polyglot, ARC-AGI-2.
Cheapest of 2 routes · OpenRouter