Models / Anthropic/ Claude Opus 5

Claude Opus 5

Anthropic · released Jul 24, 2026

Input: text, images and documents. Output: text.InputOutput
Type
Closed
Input
$5.00
Output
$25.00
Cached
$0.50

List price · per 1M tokens · Anthropic at 1M context · machine-readable source ↗

Our take

Written Aug 4, 2026

Anthropic's largest Claude model is built for demanding reasoning and coding tasks. It can handle up to one million tokens in a single request and is available across major cloud platforms. Choose it when output quality matters more than price.

Who should pick it

Pick this for your hardest reasoning, coding and agentic workloads where output quality justifies the premium price. Use it for long-document work that actually uses the one-million-token limit. Skip it if cost matters more than quality, or if you need to self-host.

The case for it

  • One-million-token limit: fits large codebases or hundreds of pages per request.
  • Same first-party pricing across five major cloud providers.
  • Accepts text, image and file inputs.

The case against it

  • Output pricing is roughly ten times that of a large downloadable model such as GLM 5.2.
  • Not downloadable, so API access only; no self-hosting path.
00

How good is it?

A closed text model for everyday questions, writing, coding and multi-step agent work.

Good at
  • getting answers to everyday questionsArena Text (overall) · 9th of 168
  • drafts, rewrites and editingArena Creative Writing · 10th of 168
  • writing and completing codeArena Coding · 9th of 168
  • multi-step work it carries out for youArena Agent · 3rd of 55
  • changing course when you give new instructionsArena Agent · Steerability · 3rd of 55
  • getting back on track after a step failsArena Agent · Recovery · 1st of 55

EverydayGeneral questions and everyday reasoning

4.5 of 5

Arena Text (overall)9th of 168 · 1491

Arena Hard Prompts 7th of 168Arena Maths 2nd of 163LiveBench Reasoning 6th of 58LiveBench Mathematics 10th of 58LiveBench Data Analysis 29th of 58

Also on this board: 1488 (Sep 25, 2026). Read the pair, not the higher one.

CodingWriting and fixing code on its own

4.5 of 5

Arena Coding9th of 168 · 1533

Arena Code (WebDev) 4th of 95LiveBench Coding 13th of 58

Also on this board: 1529 (Sep 25, 2026). Read the pair, not the higher one.

AgenticPlanning, calling tools, staying on task

4 of 5

Arena Agent3rd of 55 · 0.098

LiveBench Agentic Coding 4th of 58

Also on this board: 0.095 (Sep 25, 2026). Read the pair, not the higher one.

WritingDrafting and rewriting prose

4 of 5

Arena Creative Writing10th of 168 · 1472

LiveBench Language 4th of 58

Also on this board: 1470 (Sep 25, 2026). Read the pair, not the higher one.

How it behaves in an agent loop
Tool usereaches for the right one, and does not invent one25th of 55
Steerabilitydoes what it was asked, and changes course when told3rd of 55
Recoverygets back on track after a command fails1st of 55
Task outcomefinishes what the session set out to do5th of 55

Placings on Arena's agent boards, from live sessions people ran themselves. A model can lead on one of these and sit mid-field on the others.

Other boards it appears on
Arena Instruction Following 5th of 168LiveBench 9th of 58LiveBench Instruction Following 40th of 58Arena Agent · Recovery 1st of 55Arena Agent · Steerability 3rd of 55Arena Agent · Task outcome 5th of 55Arena Agent · Tool use 25th of 55

Boards this model appears on that none of the ratings above are built on.

Every published score for this model20 scoresEvery figure we hold, from 20 boards, with who ran it and a link to the source — including the boards no rating above is built on.
LiveBenchreasoning
80.08source ↗
65.2source ↗
81.45source ↗
74.55source ↗
63.77source ↗
88.69source ↗
95.73source ↗
91.21source ↗
0.098source ↗
0.126source ↗
0.112source ↗
0.106source ↗
0.003source ↗
1533source ↗
1472source ↗
1516source ↗
1497source ↗
1523source ↗
1491source ↗
1693source ↗
01

Where to rent it

Prices checked between 57 min and 60 min ago — each listing carries its own date.

Cheapest published offer

Anthropic, direct

The lab is also the cheapest we hold. The strip above and this offer are the same one, so nothing on this page undercuts Anthropic on 1M of context.

per 1M tokens
$5.00 in / $25.00 out
Context served
1M
Throughput
~67 tok/s
Current provider offers with price, context and prompt-privacy answers
ProviderIn / out per 1M tokensContextThroughputTrains on promptsLogs promptsZero retention
OpenRouterOpenRouter's own listing$5.00 / $25.00checked 60 min ago1Mnot measuredUnknownUnknownUnknown
DeepInfraDirect$5.00 / $25.00checked 58 min ago1Mnot measuredUnknownUnknownUnknown
Microsoft Azure AIglobalThrough OpenRouter$5.00 / $25.00checked 59 min ago1M128K max reply59 tok/sNoNoUnknown
Google Vertex AIglobalThrough OpenRouter$5.00 / $25.00checked 59 min ago1M128K max reply63 tok/sNoNoConfirmed
AnthropicDirect$5.00 / $25.00checked 57 min ago1M128K max reply67 tok/sNoYes30 daysUnknown
Amazon BedrockThrough OpenRouter$5.00 / $25.00checked 59 min ago1M128K max reply62 tok/sNoNoConfirmed
Claude Platform on AWSThrough OpenRouter$5.00 / $25.00checked 59 min ago1M128K max reply66 tok/sNoYes30 daysUnknown
Microsoft Azure AIusThrough OpenRouter$5.50 / $27.50checked 59 min ago1M128K max reply4 tok/sNoNoUnknown
Amazon Bedrockus-east-1Through OpenRouter$5.50 / $27.50checked 59 min ago1M128K max reply47 tok/sNoNoConfirmed
Amazon Bedrockeu-west-1Through OpenRouter$5.50 / $27.50checked 59 min ago1M128K max reply72 tok/sNoNoConfirmed
Google Vertex AIusThrough OpenRouter$5.50 / $27.50checked 59 min ago1M128K max reply70 tok/sNoNoConfirmed
Google Vertex AIeuropeThrough OpenRouter$5.50 / $27.50checked 59 min ago1M128K max reply65 tok/sNoNoConfirmed
Anthropicfast tierThrough OpenRouter$10.00 / $50.00checked 59 min ago1M128K max reply144 tok/sNoYes30 daysUnknown

Across the 13 listings we hold: 11 say they do not train on prompts, 0 say they do and 2 do not say. 6 appear in the zero-retention registry we check; the rest are unknown to us.

What each host's API supports

From the parameter list each endpoint publishes. Streaming is omitted: nothing we hold reports it, for any model.

API features per host
ProviderTool callingJSON outputStrict schema
OpenRouterOpenRouter's own listing✓✓✓
DeepInfraDirect
Microsoft Azure AIglobalThrough OpenRouter✓✓✓
Google Vertex AIglobalThrough OpenRouter✓✓✓
AnthropicDirect✓✓✓
Amazon BedrockThrough OpenRouter✓✓✗
Claude Platform on AWSThrough OpenRouter✓✓✓
Microsoft Azure AIusThrough OpenRouter✓✓✓
Amazon Bedrockus-east-1Through OpenRouter✓✓✗
Amazon Bedrockeu-west-1Through OpenRouter✓✓✗
Google Vertex AIusThrough OpenRouter✓✓✓
Google Vertex AIeuropeThrough OpenRouter✓✓✓
AnthropicfastThrough OpenRouter✓✓✓

Tool calling: 12 of 13 listings say yes, 1 publishes no parameter list. JSON output: 12 of 13 listings say yes, 1 publishes no parameter list. Strict schema: 9 of 13 listings say yes, 3 say no, 1 publishes no parameter list.

02

Models people weigh against Claude Opus 5

03

When we formed this view

Recent changes

Sep 25, 2026BenchmarkScored 0.098 via High on Arena Agent
What movedleaderboard
Sep 25, 2026BenchmarkScored 0.126 via Max on Arena Agent · Recovery
What movedleaderboard
Sep 25, 2026BenchmarkScored 0.112 via High on Arena Agent · Steerability
What movedleaderboard
Sep 25, 2026BenchmarkScored 0.106 via Max on Arena Agent · Task outcome
What movedleaderboard
Sep 25, 2026BenchmarkScored 0.003 via Max on Arena Agent · Tool use
What movedleaderboard
Sep 25, 2026BenchmarkScored 1533 via High on Arena Coding
What movedleaderboard
Sep 25, 2026BenchmarkScored 1472 via High on Arena Creative Writing
What movedleaderboard
Sep 25, 2026BenchmarkScored 1516 via High on Arena Hard Prompts
What movedleaderboard
Sep 25, 2026BenchmarkScored 1497 via High on Arena Instruction Following
What movedleaderboard
Sep 25, 2026BenchmarkScored 1523 via High on Arena Maths
What movedleaderboard

Each date is the day we first saw the change, or the day the maker announced it.

What we do not know about this model yet

  • 1 of 13 listings publishes no parameter list, so what its API accepts is unknown to us.
  • Nothing we hold says whether an endpoint streams, so we do not show it either way.
  • 2 of 13 listings do not say whether they train on prompts.
  • We hold no batch or off-peak rate for any of its listings.
04

Licence and identifiers

What the licence allowsWe hold no licence record for this model. Inside are the identifiers you need to pull it — its Hugging Face repo where we have one, our slug and a machine-readable card.

Licence

We hold no licence record for this model, and no record of published weights either — so we can neither summarise its terms nor point you at the weights.

Identifiers

Takes in, gives back
Text, images and documents in, text out
Catalogue slug
anthropic-claude-opus-5

Machine-readable model card (omc.json) →

Something wrong on this page? Tell us