Models / Anthropic/ Claude Opus 4.5

Claude Opus 4.5

Anthropic · released Nov 24, 2025

Input: text, images and documents. Output: text.InputOutput
Type
Proprietary
Input
$5.00
Output
$25.00
Cached
$0.50

List price · per 1M tokens · Anthropic at 200K context · source ↗

Our take

Written Aug 2, 2026

Claude Opus 4.5 is Anthropic's flagship reasoning model, available only through hosted APIs with a 200,000-token request limit. It excels at complex coding and difficult prompts, though its creative writing and maths scores sit noticeably below those peaks.

Who should pick it

Choose this for complex coding tasks where its Arena Coding Elo of 1530.2 is the headline strength, or for hard prompts and adversarial instructions at 1499.4. Use it for long-document analysis with its 200,000-token context, or when you need multimodal inputs combining text, images and files. Skip it if you are on a tight budget, need guaranteed fast throughput, or want strong creative writing from the same model.

The case for it

  • Arena Coding Elo of 1530.2, with a confirming second run at 1522.6.
  • Arena Hard Prompts Elo of 1499.4 on deliberately difficult instructions.
  • 200,000-token context window for long-form document work.
  • Consistent premium pricing across most providers, with no undercutting.

The case against it

  • Creative writing Elo of 1460.2 lags 70 points behind its own coding peak.
  • Maths Elo of 1469.3 trails its hard-prompt and coding scores by 30–61 points.
  • Premium pricing with no budget tier; throughput varies from 29 to 58 tokens per second with no guaranteed floor.
00

How good is it?

IntelligencePuzzles, maths, exam questions

4 of 5

Arena Text (overall)19th of 143 · 1468.9

Arena Hard Prompts 12th of 143Arena Maths 26th of 139LiveBench Mathematics 14th of 35 via Thinking 64k High effortLiveBench Data Analysis 15th of 35 via Thinking 64k High effortLiveBench Reasoning 23rd of 35 via Thinking 64k High effort

Also on this board: 1472.8 via Thinking 32k (Aug 2, 2026). Read the pair, not the higher one.

CodingWriting and fixing code on its own

4 of 5

Arena Coding11th of 143 · 1523

Arena Code (WebDev) 27th of 74LiveBench Coding 11th of 35 via Thinking 64k High effort

Also on this board: 1530 via Thinking 32k (Aug 2, 2026). Read the pair, not the higher one.

AgenticPlanning, calling tools, staying on task

Scored, not ratedSWE-bench Verified · 1st of 39 · 79.2via live-SWE-agent

Claude Opus 4.5 is not on Arena Agent (IPS), which is where the rating would come from, so there is no rating here. It is on SWE-bench Verified, in 1st of 39 with 79.2.

LiveBench Agentic Coding 32nd of 35 via Thinking 64k High effort

WritingWe do not rate this

Scored, not ratedThe placings are on the right.

Two boards come close and neither tests writing: Arena Creative Writing asks people which of two replies they prefer, and LiveBench Language tests whether a model understood a passage. So we show where Claude Opus 4.5 placed and give it no mark out of five.

Arena Creative Writing 13th of 143 · 1460LiveBench Language 12th of 35 · 81.3 via Thinking 64k High effort
Also scored, on boards we give no mark for
Arena Instruction Following 12th of 143LiveBench 22nd of 35 via Thinking 64k High effortLiveBench Instruction Following 26th of 35 via Thinking 64k High effort

These tests check whether a model follows instructions — a precondition for all the work above, but not a measure of how well that work is done, which is why they get no rating.

Every published score for this model16 scoresEvery figure we hold, from 16 boards, with who ran it and a link to the source — including the boards no rating above is built on.
LiveBenchreasoning
72.6via Thinking 64k High effortindependentsource ↗
39.7via Thinking 64k High effortindependentsource ↗
79.7via Thinking 64k High effortindependentsource ↗
74.4via Thinking 64k High effortindependentsource ↗
62.6via Thinking 64k High effortindependentsource ↗
81.3via Thinking 64k High effortindependentsource ↗
90.4via Thinking 64k High effortindependentsource ↗
80.1via Thinking 64k High effortindependentsource ↗
1523independentsource ↗
1497.6independentsource ↗
1465independentsource ↗
1468.9independentsource ↗
1467independentsource ↗
79.2via live-SWE-agentindependentsource ↗
01

Or rent it from someone else

Cheapest published offer

Why this differs from the header. The strip above quotes Anthropic's own list price. This is the cheapest live offer at the widest standard context we hold, whoever is serving it — a reseller undercutting a lab is ordinary commerce, not an error.

per 1M tokens
$5.00 in / $25.00 out
Context served
200K
Throughput
Not measured
Current provider offers with price, context and prompt-privacy answers
ProviderIn / out per 1M tokensContextThroughputTrains on promptsLogs promptsZero retention
OpenRouter$5.00 / $25.00200Knot measuredUnknownUnknownUnknown
Microsoft Azure AIus-east-2$5.00 / $25.00200K29 tok/sNoNoUnknown
Microsoft Azure AIglobal$5.00 / $25.00200K51 tok/sNoNoUnknown
Google Vertex AIglobal$5.00 / $25.00200K45 tok/sNoNoConfirmed
Anthropic$5.00 / $25.00200K64K out42 tok/sNoYes30 daysUnknown
Amazon Bedrock$5.00 / $25.00200K51 tok/sNoNoConfirmed
Amazon Bedrockeu-west-1$5.50 / $27.50200K3 tok/sNoNoConfirmed

Across the 7 listings we hold: 6 say they do not train on prompts, 0 say they do and 1 do not say. 3 appear in the zero-retention registry we check; the rest are unknown to us rather than confirmed either way.

What each host's API supports

From the parameter list each endpoint publishes. Streaming is omitted: nothing we hold reports it, for any model.

Supported
Not supported
Not published
host gave no parameter list
API features per host
ProviderTool callingJSON outputStrict schema
OpenRouter
Microsoft Azure AIus-east-2
Microsoft Azure AIglobal
Google Vertex AIglobal
Anthropic
Amazon Bedrock
Amazon Bedrockeu-west-1

Tool calling: 7 of 7 listings say yes. JSON output: 6 of 7 listings say yes, 1 says no. Strict schema: 6 of 7 listings say yes, 1 says no.

02

Models people weigh against Claude Opus 4.5

03

When we formed this view

Dates behind this page

Aug 2, 2026BenchmarkScored 1523 on Arena Codingleaderboard
Aug 2, 2026BenchmarkScored 1460 on Arena Creative Writingleaderboard
Aug 2, 2026BenchmarkScored 1497.6 on Arena Hard Promptsleaderboard
Aug 2, 2026BenchmarkScored 1474.3 on Arena Instruction Followingleaderboard
Aug 2, 2026BenchmarkScored 1465 on Arena Mathsleaderboard
Aug 2, 2026BenchmarkScored 1468.9 on Arena Text (overall)leaderboard
Aug 2, 2026BenchmarkScored 1467 on Arena Code (WebDev)leaderboard
Jul 27, 2026Benchmark updateclaude-opus-4-5-20251101 enters LMArena at 1469 Elo70574 votes
Jul 26, 2026ListedListed on LLMapfirst indexed by our pipeline
Jun 25, 2026BenchmarkScored 72.6 via Thinking 64k High effort on LiveBenchleaderboard

Prices last checked 8d ago

What we do not know about this model yet

  • Nothing we hold says whether an endpoint streams, so we do not show it either way.
  • 1 of 7 listings do not say whether they train on prompts.
04

Licence and identifiers

What the licence allowsWe hold no licence record for this model. Inside are the identifiers you need to pull it — its Hugging Face repo where we have one, our slug and a machine-readable card.

Licence

Commercial API terms. We hold no licence record for this model, so there is nothing to summarise here.

Identifiers

Modality record
text+image+file->text
Catalogue slug
anthropic-claude-opus-4-5

Machine-readable model card (omc.json) →

Something wrong on this page? Tell us