Models / OpenAI/ GPT-4o (2024-05-13)

GPT-4o (2024-05-13)

OpenAI · released May 13, 2024

Input: text, images and documents. Output: text.InputOutput
Type
Closed
Input
$5.00
Output
$15.00
Cached
None held

List price · per 1M tokens · OpenAI at 128K context · machine-readable source ↗

Our take

Written Sep 2, 2026

GPT-4o is a text-and-image model from OpenAI released in May 2024 that can handle up to 128,000 tokens in a single request. It carries broad measured coverage across six chat and task categories, with identical pricing across every provider channel.

Who should pick it

Pick this for general multimodal work where predictable pricing matters more than hunting for the cheapest host, or for coding workflows where its measured coding score is the relevant signal. Use the direct OpenAI channel if throughput is your bottleneck. Skip it if you need to self-host, if maths-heavy tasks dominate, or if you want measured software-engineering resolution above two-fifths.

The case for it

  • Broadest arena coverage for its vintage: six distinct categories measured, from coding to creative writing.
  • Identical pricing across all three provider channels, so provider choice is about throughput and trust, not cost.
  • Fastest measured throughput on the direct OpenAI channel, at 101 tps versus 66 tps on Azure.

The case against it

  • Maths is its weakest measured category, with a gap of over 60 points against its own coding score.
  • Resolves fewer than two in five software-engineering issues end-to-end on the benchmark we track.
  • Proprietary weights with no self-host option; parameter count undisclosed.
00

How good is it?

A general chat model for everyday questions and writing, though it trails most models at coding tasks.

Less good at
  • writing and completing codeArena Coding · 133rd of 168

EverydayGeneral questions and everyday reasoning

2 of 5

Arena Text (overall)126th of 168 · 1346

Arena Hard Prompts 132nd of 168Arena Maths 135th of 163

CodingWriting and fixing code on its own

1.5 of 5

Arena Coding133rd of 168 · 1369

Arena Coding is the only board that has scored it for this.

AgenticPlanning, calling tools, staying on task

Scored, not ratedSWE-bench Verified · 32nd of 42 · 38.8

Not yet scored on Arena Agent. It is on SWE-bench Verified, in 32nd of 42 with 38.8.

WritingDrafting and rewriting prose

2 of 5

Arena Creative Writing108th of 168 · 1338

Arena Creative Writing is the only board that has scored it for this.

Other boards it appears on
Arena Instruction Following 129th of 168

These tests check whether a model follows instructions — a precondition for all the work above, but not a measure of how well that work is done.

Every published score for this model7 scoresEvery figure we hold, from 7 boards, with who ran it and a link to the source — including the boards no rating above is built on.
1369source ↗
1338source ↗
1338source ↗
1305source ↗
1346source ↗
38.8source ↗
01

Where to rent it

Prices checked 2 hours ago — each listing carries its own date.

Cheapest published offer

OpenAI, direct

The lab is also the cheapest we hold. The strip above and this offer are the same one, so nothing on this page undercuts OpenAI on 128K of context.

per 1M tokens
$5.00 in / $15.00 out
Context served
128K
Throughput
~32 tok/s
Current provider offers with price, context and prompt-privacy answers
ProviderIn / out per 1M tokensContextThroughputTrains on promptsLogs promptsZero retention
OpenRouterOpenRouter's own listing$5.00 / $15.00checked 2 hours ago128Knot measuredUnknownUnknownUnknown
Microsoft Azure AIThrough OpenRouter$5.00 / $15.00checked 2 hours ago128K4K max reply15 tok/sNoNoConfirmed
OpenAIDirect$5.00 / $15.00checked 2 hours ago128K4K max reply32 tok/sNoYesunknown periodUnknown

Across the 3 listings we hold: 2 say they do not train on prompts, 0 say they do and 1 does not say. 1 appears in the zero-retention registry we check; the rest are unknown to us.

What each host's API supports

From the parameter list each endpoint publishes. Streaming is omitted: nothing we hold reports it, for any model.

API features per host
ProviderTool callingJSON outputStrict schema
OpenRouterOpenRouter's own listing✓✓✓
Microsoft Azure AIThrough OpenRouter✓✓✓
OpenAIDirect✓✓✓

Tool calling: 3 of 3 listings say yes. JSON output: 3 of 3 listings say yes. Strict schema: 3 of 3 listings say yes.

02

Models people weigh against GPT-4o (2024-05-13)

03

When we formed this view

Recent changes

Sep 25, 2026BenchmarkScored 1369 on Arena Coding
What movedleaderboard
Sep 25, 2026BenchmarkScored 1338 on Arena Creative Writing
What movedleaderboard
Sep 25, 2026BenchmarkScored 1338 on Arena Hard Prompts
What movedleaderboard
Sep 25, 2026BenchmarkScored 1326 on Arena Instruction Following
What movedleaderboard
Sep 25, 2026BenchmarkScored 1305 on Arena Maths
What movedleaderboard
Sep 25, 2026BenchmarkScored 1346 on Arena Text (overall)
What movedleaderboard
Jul 26, 2026ListedListed on LLMap
What movedfirst indexed by our pipeline
Oct 28, 2024BenchmarkScored 38.8 via Agentless-1.5 on SWE-bench Verified
What movedleaderboard
May 13, 2024AnnouncedGPT-4o (2024-05-13) announced by OpenAI

Each date is the day we first saw the change, or the day the maker announced it.

What we do not know about this model yet

  • Nothing we hold says whether an endpoint streams, so we do not show it either way.
  • 1 of 3 listings does not say whether it trains on prompts.
  • We hold no cached-input rate for any of its listings.
  • We hold no batch or off-peak rate for any of its listings.
04

Licence and identifiers

What the licence allowsWe hold no licence record for this model. Inside are the identifiers you need to pull it — its Hugging Face repo where we have one, our slug and a machine-readable card.

Licence

We hold no licence record for this model, and no record of published weights either — so we can neither summarise its terms nor point you at the weights.

Identifiers

Takes in, gives back
Text, images and documents in, text out
Catalogue slug
openai-gpt-4o-2024-05-13

Machine-readable model card (omc.json) →

Something wrong on this page? Tell us