Models / OpenAI/ GPT-5.1-Codex-Mini

GPT-5.1-Codex-Mini

OpenAI · released Nov 13, 2025

Input: text and images. Output: text.InputOutput
Type
Proprietary
Input
None held
Output
None held
Cached
None held

We don't hold a list price for this model yet · hosted only — we hold no downloadable copy

Our take

Written Aug 3, 2026

GPT-5.1-Codex-Mini is a compact coding specialist from OpenAI that can handle up to 400,000 tokens in a single request and accepts both text and images. It is built for long-context coding workflows, particularly in web development, though its measured quality rests on a single benchmark type.

Who should pick it

Pick this for long-context coding workflows where you need 400,000 tokens of working memory, or for web development tasks where the Arena Code (WebDev) leaderboard is your relevant quality signal. Use it on Azure if you need guaranteed throughput at 77 tokens per second. Skip it if you need evidence of general chat, reasoning, or non-web coding ability, or if you want to run the model on your own hardware.

The case for it

  • Extremely long context for a coding model: up to 400,000 tokens in a single request.
  • Stable leaderboard presence with tight score clustering: six consecutive daily readings within a spread of only 0.1215 points.

The case against it

  • Thin benchmark coverage — only one task type measured, with no general chat, reasoning, or multilingual coding scores.
  • No throughput data on OpenRouter, and modest speed on Azure at 77 tokens per second.
  • Output costs several times more than input, with no cheaper tier available.
00

How good is it?

IntelligencePuzzles, maths, exam questions

not measured

Nobody we watch has scored GPT-5.1-Codex-Mini for this. We would take the rating from Arena Text (overall).

CodingWriting and fixing code on its own

Scored, not ratedArena Code (WebDev) · 67th of 74 · 1244.2

GPT-5.1-Codex-Mini is not on Arena Coding, which is where the rating would come from, so there is no rating here. It is on Arena Code (WebDev), in 67th of 74 with 1244.2.

AgenticPlanning, calling tools, staying on task

not measured

Nobody we watch has scored GPT-5.1-Codex-Mini for this. We would take the rating from Arena Agent (IPS).

WritingWe do not rate this

not measured

Nobody we watch has scored this model for writing. Two boards come close and neither tests writing: Arena Creative Writing asks people which of two replies they prefer, and LiveBench Language tests whether a model understood a passage.

Every published score for this model1 scoreEvery figure we hold, from 1 board, with who ran it and a link to the source — including the boards no rating above is built on.
1244.2independentsource ↗
01

Or rent it from someone else

Cheapest published offer

Cheapest of 2 live listings. Picked at the widest standard context we hold, within one quantisation slice, so the numbers beside it are a price one host actually charges.

per 1M tokens
$0.25 in / $2.00 out
Context served
400K
Throughput
Not measured
Current provider offers with price, context and prompt-privacy answers
ProviderIn / out per 1M tokensContextThroughputTrains on promptsLogs promptsZero retention
OpenRouter$0.25 / $2.00400Knot measuredUnknownUnknownUnknown
Microsoft Azure AI$0.25 / $2.00400K39 tok/sNoNoConfirmed

Across the 2 listings we hold: 1 say they do not train on prompts, 0 say they do and 1 do not say. 1 appear in the zero-retention registry we check; the rest are unknown to us rather than confirmed either way.

What each host's API supports

From the parameter list each endpoint publishes. Streaming is omitted: nothing we hold reports it, for any model.

Supported
Not supported
Not published
host gave no parameter list
API features per host
ProviderTool callingJSON outputStrict schema
OpenRouter
Microsoft Azure AI

Tool calling: 2 of 2 listings say yes. JSON output: 2 of 2 listings say yes. Strict schema: 2 of 2 listings say yes.

02

When we formed this view

Dates behind this page

Aug 2, 2026BenchmarkScored 1244.2 on Arena Code (WebDev)leaderboard
Jul 26, 2026ListedListed on LLMapfirst indexed by our pipeline
Nov 13, 2025AnnouncedGPT-5.1-Codex-Mini announced by OpenAI

Prices last checked 9d ago

What we do not know about this model yet

  • Nothing we hold says whether an endpoint streams, so we do not show it either way.
  • 1 of 2 listings do not say whether they train on prompts.
  • We don't hold a list price for this model yet — the gap is ours, not the lab's.
03

Licence and identifiers

What the licence allowsWe hold no licence record for this model. Inside are the identifiers you need to pull it — its Hugging Face repo where we have one, our slug and a machine-readable card.

Licence

Commercial API terms. We hold no licence record for this model, so there is nothing to summarise here.

Identifiers

Modality record
text+image->text
Catalogue slug
openai-gpt-5-1-codex-mini

Machine-readable model card (omc.json) →

Something wrong on this page? Tell us