GPT-4o (2024-05-13)
OpenAI · released May 13, 2024
- Type
- Closed
- Input
- $5.00
- Output
- $15.00
- Cached
- None held
List price · per 1M tokens · OpenAI at 128K context · machine-readable source ↗
Our take
Written Sep 2, 2026GPT-4o is a text-and-image model from OpenAI released in May 2024 that can handle up to 128,000 tokens in a single request. It carries broad measured coverage across six chat and task categories, with identical pricing across every provider channel.
Pick this for general multimodal work where predictable pricing matters more than hunting for the cheapest host, or for coding workflows where its measured coding score is the relevant signal. Use the direct OpenAI channel if throughput is your bottleneck. Skip it if you need to self-host, if maths-heavy tasks dominate, or if you want measured software-engineering resolution above two-fifths.
The case for it
- Broadest arena coverage for its vintage: six distinct categories measured, from coding to creative writing.
- Identical pricing across all three provider channels, so provider choice is about throughput and trust, not cost.
- Fastest measured throughput on the direct OpenAI channel, at 101 tps versus 66 tps on Azure.
The case against it
- Maths is its weakest measured category, with a gap of over 60 points against its own coding score.
- Resolves fewer than two in five software-engineering issues end-to-end on the benchmark we track.
- Proprietary weights with no self-host option; parameter count undisclosed.
How good is it?
A general chat model for everyday questions and writing, though it trails most models at coding tasks.
- writing and completing codeArena Coding · 133rd of 168
EverydayGeneral questions and everyday reasoning
Arena Text (overall)126th of 168 · 1346
CodingWriting and fixing code on its own
Arena Coding133rd of 168 · 1369
Arena Coding is the only board that has scored it for this.
AgenticPlanning, calling tools, staying on task
Not yet scored on Arena Agent. It is on SWE-bench Verified, in 32nd of 42 with 38.8.
WritingDrafting and rewriting prose
Arena Creative Writing108th of 168 · 1338
Arena Creative Writing is the only board that has scored it for this.
These tests check whether a model follows instructions — a precondition for all the work above, but not a measure of how well that work is done.
Every published score for this model7 scoresEvery figure we hold, from 7 boards, with who ran it and a link to the source — including the boards no rating above is built on.
Where to rent it
Prices checked 4 hours ago — each listing carries its own date.
OpenAI, direct
The lab is also the cheapest we hold. The strip above and this offer are the same one, so nothing on this page undercuts OpenAI on 128K of context.
- per 1M tokens
- $5.00 in / $15.00 out
- Context served
- 128K
- Throughput
- ~32 tok/s
| Provider | In / out per 1M tokens | Context | Throughput | Trains on prompts | Logs prompts | Zero retention |
|---|---|---|---|---|---|---|
| OpenRouterOpenRouter's own listing | $5.00 / $15.00checked 4 hours ago | 128K | not measured | Unknown | Unknown | Unknown |
| Microsoft Azure AIThrough OpenRouter | $5.00 / $15.00checked 4 hours ago | 128K4K max reply | 15 tok/s | No | No | Confirmed |
| OpenAIDirect | $5.00 / $15.00checked 4 hours ago | 128K4K max reply | 32 tok/s | No | Yesunknown period | Unknown |
Across the 3 listings we hold: 2 say they do not train on prompts, 0 say they do and 1 does not say. 1 appears in the zero-retention registry we check; the rest are unknown to us.
What each host's API supports
From the parameter list each endpoint publishes. Streaming is omitted: nothing we hold reports it, for any model.
| Provider | Tool calling | JSON output | Strict schema |
|---|---|---|---|
| OpenRouterOpenRouter's own listing | ✓ | ✓ | ✓ |
| Microsoft Azure AIThrough OpenRouter | ✓ | ✓ | ✓ |
| OpenAIDirect | ✓ | ✓ | ✓ |
Tool calling: 3 of 3 listings say yes. JSON output: 3 of 3 listings say yes. Strict schema: 3 of 3 listings say yes.
Models people weigh against GPT-4o (2024-05-13)
When we formed this view
Recent changes
What moved
first indexed by our pipelineEach date is the day we first saw the change, or the day the maker announced it.
What we do not know about this model yet
- Nothing we hold says whether an endpoint streams, so we do not show it either way.
- 1 of 3 listings does not say whether it trains on prompts.
- We hold no cached-input rate for any of its listings.
- We hold no batch or off-peak rate for any of its listings.
Licence and identifiers
What the licence allowsWe hold no licence record for this model. Inside are the identifiers you need to pull it — its Hugging Face repo where we have one, our slug and a machine-readable card.
Licence
We hold no licence record for this model, and no record of published weights either — so we can neither summarise its terms nor point you at the weights.
Identifiers
- Takes in, gives back
- Text, images and documents in, text out
- Catalogue slug
- openai-gpt-4o-2024-05-13