Claude Sonnet 4
Anthropic · released May 22, 2025
- Type
- Closed
- Input
- None held
- Output
- None held
- Cached
- None held
We don't hold a list price for this model yet · hosted only — we have no record of published weights
Our take
Written Sep 29, 2026Claude Sonnet 4 is a hosted-only model from Anthropic: we list no download for it, so using it means choosing a host. It takes text, images and files in one request and sits mid-field on the Arena boards, so treat it as a solid generalist rather than a leader.
Reach for it when a single request has to carry text, images and files together, or when long documents would otherwise need splitting up first. It suits teams that want a hosted model at one flat rate across the offers we list rather than running anything themselves. Skip it if you need a model near the top of the Arena boards, or if you need to run the model on your own hardware.
The case for it
- Text, images and files go into the same request, so a screenshot or a document does not have to be described in words first.
- The request capacity takes a long report or a stack of documents beside the question, though reliable recall across all of it is unverified in our data.
- 64.9% of real GitHub issues resolved end-to-end on SWE-bench Verified, 19th of 42 as of 19 Feb 2026 — a set-piece benchmark of issue resolution, not a measure of working inside your own repository.
The case against it
- 105th of 168 on Arena Text (overall) as of 25 Sep 2026, with 95th of 168 on Arena Coding and 105th of 163 on Arena Maths the same day — these boards record which answer people preferred, not whether it was correct.
- 59.4% pass on contamination-free competitive programming problems on LiveCodeBench, 10th of 16 as of 28 Sep 2026 — set-piece exercises rather than fixing issues in an existing project.
- We list no download for it, so using it means choosing a host, and no licence is supplied for the weights.
How good is it?
EverydayGeneral questions and everyday reasoning
Arena Text (overall)105th of 168 · 1391
Also on this board: 1402 (Sep 25, 2026). Read the pair, not the higher one.
CodingWriting and fixing code on its own
Arena Coding95th of 168 · 1449
Also on this board: 1474 (Sep 25, 2026). Read the pair, not the higher one.
AgenticPlanning, calling tools, staying on task
Not yet scored on Arena Agent. It is on SWE-bench Verified, in 19th of 42 with 64.9.
WritingDrafting and rewriting prose
Arena Creative Writing80th of 168 · 1388
Arena Creative Writing is the only board that has scored it for this.
Also on this board: 1395 (Sep 25, 2026). Read the pair, not the higher one.
These tests check whether a model follows instructions — a precondition for all the work above, but not a measure of how well that work is done.
Every published score for this model8 scoresEvery figure we hold, from 8 boards, with who ran it and a link to the source — including the boards no rating above is built on.
Where to rent it
Prices checked 58 min ago — each listing carries its own date.
- per 1M tokens
- $3.00 in / $15.00 out
- Context served
- 200K
- Throughput
- ~43 tok/s
| Provider | In / out per 1M tokens | Context | Throughput | Trains on prompts | Logs prompts | Zero retention |
|---|---|---|---|---|---|---|
| OpenRouterOpenRouter's own listing | $3.00 / $15.00checked 58 min ago | 200K | not measured | Unknown | Unknown | Unknown |
| Amazon BedrockThrough OpenRouter | $3.00 / $15.00checked 58 min ago | 200K64K max reply | 43 tok/s | No | No | Confirmed |
| Amazon Bedrockeu-west-1Through OpenRouter | $3.00 / $15.00checked 58 min ago | 200K64K max reply | 3 tok/s | No | No | Confirmed |
Across the 3 listings we hold: 2 say they do not train on prompts, 0 say they do and 1 does not say. 2 appear in the zero-retention registry we check; the rest are unknown to us.
What each host's API supports
From the parameter list each endpoint publishes. Streaming is omitted: nothing we hold reports it, for any model.
| Provider | Tool calling | JSON output | Strict schema |
|---|---|---|---|
| OpenRouterOpenRouter's own listing | ✓ | ✗ | ✗ |
| Amazon BedrockThrough OpenRouter | ✓ | ✗ | ✗ |
| Amazon Bedrockeu-west-1Through OpenRouter | ✓ | ✗ | ✗ |
Tool calling: 3 of 3 listings say yes. JSON output: 0 of 3 listings say yes, 3 say no. Strict schema: 0 of 3 listings say yes, 3 say no.
Models people weigh against Claude Sonnet 4
When we formed this view
Recent changes
What moved
first indexed by our pipelineEach date is the day we first saw the change, or the day the maker announced it.
What we do not know about this model yet
- Nothing we hold says whether an endpoint streams, so we do not show it either way.
- 1 of 3 listings does not say whether it trains on prompts.
- We don't hold a list price for this model yet — the gap is ours, not the lab's.
- We hold no batch or off-peak rate for any of its listings.
Licence and identifiers
What the licence allowsWe hold no licence record for this model. Inside are the identifiers you need to pull it — its Hugging Face repo where we have one, our slug and a machine-readable card.
Licence
We hold no licence record for this model, and no record of published weights either — so we can neither summarise its terms nor point you at the weights.
Identifiers
- Takes in, gives back
- Text, images and documents in, text out
- Catalogue slug
- anthropic-claude-sonnet-4