Models / xAI/ Grok 4.5

Grok 4.5

xAI · released Jul 8, 2026

Input: text, images and documents. Output: text.InputOutput
Type
Closed
Input
$2.00
Output
$6.00
Cached
None held

List price · per 1M tokens · xAI at 500K context · machine-readable source ↗

Our take

Written Aug 4, 2026

xAI's top-tier chat model placed second on Arena Text, the independent human-preference board, in July 2026. Its output price is lower than typical for the tier, but its request limit is half the size of the one-million-token peers.

Who should pick it

Pick this for chat-quality-sensitive products on a budget. Use it if your context needs fit within 500,000 tokens. Skip it if you need a one-million-token context or multi-cloud hosting.

The case for it

  • A first-tier Arena Text placing (July 2026) at a mid-tier output price.
  • Low base price for the tier.

The case against it

  • 500,000-token request limit, half the size of the one-million-token peers.
  • Only xAI first-party and OpenRouter in our data, so fewer hosting routes than Claude, Gemini or the GPT line.
00

How good is it?

A closed text model for everyday questions, drafting, coding and tool use, including picking itself up after a failed step.

Good at
  • getting answers to everyday questionsArena Text (overall) · 32nd of 168
  • drafts, rewrites and editingArena Creative Writing · 24th of 168
  • writing and completing codeArena Coding · 33rd of 168
  • calling tools to carry out requestsArena Agent · Tool use · 9th of 55
  • getting back on track after a step failsArena Agent · Recovery · 13th of 55

EverydayGeneral questions and everyday reasoning

4 of 5

Arena Text (overall)32nd of 168 · 1465

Arena Hard Prompts 30th of 168Arena Maths 32nd of 163LiveBench Reasoning 26th of 58LiveBench Mathematics 27th of 58LiveBench Data Analysis 36th of 58

CodingWriting and fixing code on its own

4 of 5

Arena Coding33rd of 168 · 1513

Arena Code (WebDev) 26th of 95LiveBench Coding 55th of 58

AgenticPlanning, calling tools, staying on task

2.5 of 5

Arena Agent22nd of 55 · 0.016

LiveBench Agentic Coding 21st of 58

WritingDrafting and rewriting prose

3.5 of 5

Arena Creative Writing24th of 168 · 1452

LiveBench Language 19th of 58
How it behaves in an agent loop
Tool usereaches for the right one, and does not invent one9th of 55
Steerabilitydoes what it was asked, and changes course when told14th of 55
Recoverygets back on track after a command fails13th of 55
Task outcomefinishes what the session set out to do27th of 55

Placings on Arena's agent boards, from live sessions people ran themselves. A model can lead on one of these and sit mid-field on the others.

Other boards it appears on
Arena Instruction Following 29th of 168LiveBench Instruction Following 20th of 58LiveBench 29th of 58Arena Agent · Tool use 9th of 55Arena Agent · Recovery 13th of 55Arena Agent · Steerability 14th of 55Arena Agent · Task outcome 27th of 55

Boards this model appears on that none of the ratings above are built on.

Every published score for this model20 scoresEvery figure we hold, from 20 boards, with who ran it and a link to the source — including the boards no rating above is built on.
LiveBenchreasoning
75.77source ↗
56.46source ↗
68.59source ↗
73.04source ↗
82.8source ↗
90.82source ↗
87.17source ↗
0.016source ↗
0.061source ↗
0.043source ↗
−0.006source ↗
0.004source ↗
1513source ↗
1452source ↗
1489source ↗
1471source ↗
1465source ↗
1552source ↗
01

Where to rent it

Prices checked between 59 min and 1 hour ago — each listing carries its own date.

Cheapest published offer

xAI, direct and through OpenRouter

The lab is also the cheapest we hold. The strip above and this offer are the same one, so nothing on this page undercuts xAI on 500K of context.

per 1M tokens
$2.00 in / $6.00 out
Context served
500K
Throughput
Not measured
Current provider offers with price, context and prompt-privacy answers
ProviderIn / out per 1M tokensContextThroughputTrains on promptsLogs promptsZero retention
OpenRouterOpenRouter's own listing$2.00 / $6.00checked 1 hour ago500Knot measuredUnknownUnknownUnknown
xAIDirect and through OpenRouter$2.00 / $6.00checked 59 min ago directchecked 1 hour ago through OpenRouter500K500K max reply direct450K max reply through OpenRouter61 tok/sthrough OpenRouterDirectUnknownThrough OpenRouterNoDirectUnknownThrough OpenRouterYes30 daysDirectUnknownThrough OpenRouterConfirmed
xAIpriority tierThrough OpenRouter$4.00 / $12.00checked 1 hour ago500K450K max reply66 tok/sNoYes30 daysConfirmed

Across the 3 listings we hold: 2 say they do not train on prompts (1 of them only through OpenRouter), 0 say they do and 1 does not say. 2 appear in the zero-retention registry we check (1 of them only through OpenRouter); the rest are unknown to us.

What each host's API supports

From the parameter list each endpoint publishes. Streaming is omitted: nothing we hold reports it, for any model.

API features per host
ProviderTool callingJSON outputStrict schema
OpenRouterOpenRouter's own listing✓✓✓
xAIDirect and through OpenRouter✓✓✓
xAIpriorityThrough OpenRouter✓✓✓

Tool calling: 3 of 3 listings say yes. JSON output: 3 of 3 listings say yes. Strict schema: 3 of 3 listings say yes.

02

Models people weigh against Grok 4.5

03

When we formed this view

Recent changes

Sep 25, 2026BenchmarkScored 0.016 on Arena Agent
What movedleaderboard
Sep 25, 2026BenchmarkScored 0.061 on Arena Agent · Recovery
What movedleaderboard
Sep 25, 2026BenchmarkScored 0.043 on Arena Agent · Steerability
What movedleaderboard
Sep 25, 2026BenchmarkScored −0.006 on Arena Agent · Task outcome
What movedleaderboard
Sep 25, 2026BenchmarkScored 0.004 on Arena Agent · Tool use
What movedleaderboard
Sep 25, 2026BenchmarkScored 1513 on Arena Coding
What movedleaderboard
Sep 25, 2026BenchmarkScored 1452 on Arena Creative Writing
What movedleaderboard
Sep 25, 2026BenchmarkScored 1489 on Arena Hard Prompts
What movedleaderboard
Sep 25, 2026BenchmarkScored 1462 on Arena Instruction Following
What movedleaderboard
Sep 25, 2026BenchmarkScored 1471 on Arena Maths
What movedleaderboard

Each date is the day we first saw the change, or the day the maker announced it.

What we do not know about this model yet

  • Nothing we hold says whether an endpoint streams, so we do not show it either way.
  • 1 of 3 listings does not say whether it trains on prompts, and 1 answers only through OpenRouter, not for its own listing.
  • We hold no batch or off-peak rate for any of its listings.
04

Licence and identifiers

What the licence allowsWe hold no licence record for this model. Inside are the identifiers you need to pull it — its Hugging Face repo where we have one, our slug and a machine-readable card.

Licence

We hold no licence record for this model, and no record of published weights either — so we can neither summarise its terms nor point you at the weights.

Identifiers

Takes in, gives back
Text, images and documents in, text out
Catalogue slug
x-ai-grok-4-5

Machine-readable model card (omc.json) →

Something wrong on this page? Tell us