Models / OpenAI/ GPT-5.4 Mini

GPT-5.4 Mini

OpenAI · released Mar 17, 2026

Input: text, images and documents. Output: text.InputOutput
Type
Closed
Input
$0.75
Output
$4.50
Cached
$0.075

List price · per 1M tokens · OpenAI at 272K context · machine-readable source ↗

Our take

Written Sep 30, 2026

GPT-5.4 Mini is a hosted-only model: we list no download for it, so using it means choosing a host. Its measured results are mid-pack on the Arena boards, with its strongest showing on set-piece coding and maths tasks and its weakest inside an agent harness.

Who should pick it

Reach for it on set-piece code generation and completion, or on competition and olympiad-style maths, where its LiveBench figures are the strongest of its measured categories. It takes text, images and files, and its request capacity is large enough that a long report need not be split up first. Skip it if the job is agentic coding inside a harness, or if you meant to run the model on your own machine.

The case for it

  • Competition and olympiad maths is its strongest measured category, at 78.46% on LiveBench with the xHigh variant.
  • Code generation and completion follows at 71.62% on LiveBench with the xHigh variant, so set-piece coding work is where to start.
  • A long report or a stack of documents fits beside the question, though reliable recall across all of it is unverified in our data.
  • Images and files go into the same request as the text, so a screenshot or a document does not have to be described in words first.

The case against it

  • Agentic coding inside a harness is its weakest measured area, at 41.67% on LiveBench with the xHigh variant.
  • Mid-pack on the Arena boards: 52nd of 168 on Arena Text (overall) via High as of 25 Sep 2026, and 67th of 168 on Arena Creative Writing via High as of 25 Sep 2026.
  • We list no download for it, so using it means choosing a host.
00

How good is it?

EverydayGeneral questions and everyday reasoning

3 of 5

Arena Text (overall)52nd of 168 · 1447

Arena Hard Prompts 52nd of 168Arena Maths 57th of 163LiveBench Data Analysis 41st of 58LiveBench Mathematics 54th of 58LiveBench Reasoning 54th of 58

CodingWriting and fixing code on its own

3.5 of 5

Arena Coding57th of 168 · 1495

Arena Code (WebDev) 62nd of 95LiveBench Coding 48th of 58

AgenticPlanning, calling tools, staying on task

Scored, not ratedLiveBench Agentic Coding · 51st of 58 · 41.67

Not yet scored on Arena Agent. It is on LiveBench Agentic Coding, in 51st of 58 with 41.67.

WritingDrafting and rewriting prose

2.5 of 5

Arena Creative Writing67th of 168 · 1402

LiveBench Language 53rd of 58
Other boards it appears on
Arena Instruction Following 59th of 168LiveBench Instruction Following 51st of 58LiveBench 54th of 58

Boards this model appears on that none of the ratings above are built on.

Every published score for this model15 scoresEvery figure we hold, from 15 boards, with who ran it and a link to the source — including the boards no rating above is built on.
LiveBenchreasoning
66.37source ↗
41.67source ↗
71.62source ↗
70.79source ↗
59.8source ↗
70.95source ↗
78.46source ↗
71.32source ↗
1495source ↗
1402source ↗
1469source ↗
1434source ↗
1439source ↗
1447source ↗
1397source ↗
01

Where to rent it

Prices checked between 59 min and 1 hour ago — each listing carries its own date.

Cheapest published offer

Microsoft Azure AI, through OpenRouter

Why this differs from the header. The strip above quotes OpenAI's own list price; this is the cheapest live offer, whoever is serving it — a reseller undercutting a lab is ordinary commerce, not an error.

per 1M tokens
$0.75 in / $4.50 out
Context served
400K
Throughput
~46 tok/s
Current provider offers with price, context and prompt-privacy answers
ProviderIn / out per 1M tokensContextThroughputTrains on promptsLogs promptsZero retention
OpenAIflex tierThrough OpenRouter$0.38 / $2.25checked 1 hour ago400K128K max reply78 tok/sNoYesunknown periodUnknown
OpenRouterOpenRouter's own listing$0.75 / $4.50checked 1 hour ago400Knot measuredUnknownUnknownUnknown
Microsoft Azure AIThrough OpenRouter$0.75 / $4.50checked 1 hour ago400K128K max reply46 tok/sNoNoConfirmed
OpenAIDirect$0.75 / $4.50checked 59 min ago272K128K max reply66 tok/sNoYesunknown periodUnknown
Microsoft Azure AIusThrough OpenRouter$0.82 / $4.95checked 1 hour ago400K128K max reply25 tok/sNoNoConfirmed
OpenAIfast tierThrough OpenRouter$1.50 / $9.00checked 1 hour ago400K128K max reply67 tok/sNoYesunknown periodUnknown

Across the 6 listings we hold: 5 say they do not train on prompts, 0 say they do and 1 does not say. 2 appear in the zero-retention registry we check; the rest are unknown to us.

What each host's API supports

From the parameter list each endpoint publishes. Streaming is omitted: nothing we hold reports it, for any model.

API features per host
ProviderTool callingJSON outputStrict schema
OpenAIflexThrough OpenRouter✓✓✓
OpenRouterOpenRouter's own listing✓✓✓
Microsoft Azure AIThrough OpenRouter✓✓✓
OpenAIDirect✓✓✓
Microsoft Azure AIusThrough OpenRouter✓✓✓
OpenAIfastThrough OpenRouter✓✓✓

Tool calling: 6 of 6 listings say yes. JSON output: 6 of 6 listings say yes. Strict schema: 6 of 6 listings say yes.

02

Models people weigh against GPT-5.4 Mini

03

When we formed this view

Recent changes

Sep 25, 2026BenchmarkScored 1495 via High on Arena Coding
What movedleaderboard
Sep 25, 2026BenchmarkScored 1402 via High on Arena Creative Writing
What movedleaderboard
Sep 25, 2026BenchmarkScored 1469 via High on Arena Hard Prompts
What movedleaderboard
Sep 25, 2026BenchmarkScored 1434 via High on Arena Instruction Following
What movedleaderboard
Sep 25, 2026BenchmarkScored 1439 via High on Arena Maths
What movedleaderboard
Sep 25, 2026BenchmarkScored 1447 via High on Arena Text (overall)
What movedleaderboard
Sep 25, 2026BenchmarkScored 1397 via High on Arena Code (WebDev)
What movedleaderboard
Jul 26, 2026ListedListed on LLMap
What movedfirst indexed by our pipeline
Jun 25, 2026BenchmarkScored 66.37 via xHigh on LiveBench
What movedleaderboard
Jun 25, 2026BenchmarkScored 41.67 via xHigh on LiveBench Agentic Coding
What movedleaderboard

Each date is the day we first saw the change, or the day the maker announced it.

What we do not know about this model yet

  • Nothing we hold says whether an endpoint streams, so we do not show it either way.
  • 1 of 6 listings does not say whether it trains on prompts.
  • We hold no batch or off-peak rate for any of its listings.
04

Licence and identifiers

What the licence allowsWe hold no licence record for this model. Inside are the identifiers you need to pull it — its Hugging Face repo where we have one, our slug and a machine-readable card.

Licence

We hold no licence record for this model, and no record of published weights either — so we can neither summarise its terms nor point you at the weights.

Identifiers

Takes in, gives back
Text, images and documents in, text out
Catalogue slug
openai-gpt-5-4-mini

Machine-readable model card (omc.json) →

Something wrong on this page? Tell us