Grok 4.5 vs Grok 4.6
Benchmark Performance
Available Benchmarks
| Benchmark | Grok 4.5 | Grok 4.6 |
|---|---|---|
| ARC-AGI-1verified-v1-15fb467fd4fc · verified_score · leader | 87.17100% of row best · percent · Grok 4.5 (Medium) | 87.50100% of row best · percent · Grok 4.6 (Medium) |
| ARC-AGI-2verified-v2-9a3db289984e · verified_score · leader | 52.6478% of row best · percent · Grok 4.5 (Medium) | 67.08100% of row best · percent · Grok 4.6 (XHigh) |
| LMArena Agent Arenaagent-2026-08-31-011508720696 · outcome_score · statistical tie | 6.06100% of row best · score · Grok 4.5; 95% CI [4.91016619, 7.21148996]; sessions 34008; observations 1655564; rank 13 | 5.76100% of row best · score · Grok 4.6 (xHigh); 95% CI [4.47058991, 7.04827601]; sessions 15090; observations 1574671; rank 16 |
| LMArena Text Arenatext-2026-09-01-011508720696 · arena_rating · statistical tie | 1,451.82100% of row best · rating · grok-4.5; 95% CI [1446.90989155, 1456.73122428]; votes 25873; rank 39 | 1,443.7599% of row best · rating · grok-4.6-high; 95% CI [1433.62910542, 1453.86294509]; votes 3453; rank 55 |
| LiveBench2026-06-25 · overall · leader | 80.0197% of row best · percent · grok-4.5 · 10,517 output tokens / case | 82.17100% of row best · percent · grok-4.6 · 16,154 output tokens / case |
| ToneBench2026-08-28-10-task-cd9819ab6e4d · overall_score · leader | 82.2395% of row best · points · Grok 4.5 · 2,688 output tokens / case | 86.61100% of row best · points · Grok 4.6 · 15,575 output tokens / case |
| Overall ResultCounted from the protocol-matched rows above · 2 ties | 0 benchmark wins | 4 benchmark winsOverall lead |
Third-party benchmark Only like-for-like primary-publisher results are shown; raw scores, relative scores, configuration, and token spend remain visible.
Technical Differences
Side-by-Side Facts
| Field | Grok 4.5 | Grok 4.6 |
|---|---|---|
| Developer | Xai | Xai |
| Family | Grok 4 | Grok 4 |
| Model | Grok 4.5 | Grok 4.6 |
| Version | Grok 4.5 | Grok 4.6 |
| Lifecycle | active | active |
| Released | 2026-07-16 | Unknown |
| Knowledge cutoff | Unknown | 2026-02-01 |
| Input modalities | Unknown | Unknown |
| Output modalities | Unknown | Unknown |
| Context window | 500,000 | 500,000 |
| Total parameters | Unknown | Unknown |
| Active parameters | Unknown | Unknown |
| License | Unknown | Unknown |
| Open weights | No | No |
| API available | Yes | Yes |
| Self-hostable | No | No |
| Provider access | Xai (Standard) | Xai (Standard) |
| Capabilities | chat, generation, reasoning, structured_outputs, tools | chat, generation, reasoning, tools |
11 comparable fields · 3 material differences · Interactive comparison only; indexing gate not met
Grok 4.5 Capabilities
Grok 4.6 Capabilities
Primary Evidence
Sources and Freshness
Questions
Grok 4.5 vs Grok 4.6 FAQs
Is Grok 4.5 or Grok 4.6 better for coding?+
This comparison does not currently contain a protocol-matched coding benchmark for both Grok 4.5 and Grok 4.6, so Model Markets cannot name a coding leader from pricing, context size, or capability labels alone.
Which is cheaper, Grok 4.5 or Grok 4.6?+
Neither model has a directly sourced input price in this comparison. Neither model has a directly sourced output price in this comparison.
Which has a larger context window, Grok 4.5 or Grok 4.6?+
Neither model has a larger sourced context window in this comparison. Grok 4.5 is 500,000 and Grok 4.6 is 500,000.
Which performs better in benchmarks, Grok 4.5 or Grok 4.6?+
Grok 4.6 leads the current overall benchmark count. The result uses 6 protocol-matched benchmarks from 4 publishers; it is not a universal quality score.
Can Grok 4.5 or Grok 4.6 be self-hosted?+
Both models have the same recorded self-hosting status: unsupported. Grok 4.5 is not marked open weight; Grok 4.6 is not marked open weight.
Can Grok 4.5 and Grok 4.6 understand images?+
Grok 4.5 is not documented with image input; Grok 4.6 is not documented with image input. This reflects supported input modalities, not vision quality.
Which can generate longer answers, Grok 4.5 or Grok 4.6?+
Neither has a larger sourced maximum output. Grok 4.5 is — and Grok 4.6 is —.
Do Grok 4.5 and Grok 4.6 support reasoning and tool use?+
Grok 4.5: reasoning and tool calling. Grok 4.6: reasoning and tool calling. Feature support does not establish relative quality.
Which is available from more inference providers, Grok 4.5 or Grok 4.6?+
Grok 4.5 has 1 sourced provider route; Grok 4.6 has 1, a tie.
Which offers better value, Grok 4.5 or Grok 4.6?+
There is no universal value winner. Compare the input and output prices above with the matched benchmark result for your workload: cheaper tokens can be offset by different quality, token usage, latency, or provider availability.