Models / Anthropic/ Claude Opus 4.6

Claude Opus 4.6

Anthropic · released Feb 4, 2026

Input: text, images and documents. Output: text.InputOutput
Type
Proprietary
Input
$5.00
Output
$25.00
Cached
$0.50

List price · per 1M tokens · Anthropic at 1M context · source ↗

Our take

Written Aug 3, 2026

Claude Opus 4.6 is Anthropic's flagship hosted-only reasoning model with a one-million-token request limit and top scores on the independent coding and hard-prompt leaderboards we track. It is built for demanding long-document and coding work, but costs noticeably more than most alternatives and cannot be downloaded or fine-tuned.

Who should pick it

Pick this for highest-stakes coding and reasoning where its leaderboard scores justify the premium, or for very long documents needing the widest context window we track. Use it if your enterprise is already on Anthropic, Amazon Bedrock, Azure, OpenRouter or Google Vertex. Skip it if budget matters more than marginal quality gains, if you need consistent throughput, or if you require downloadable weights.

The case for it

  • One-million-token request limit, the widest in our data.
  • Top-quartile coding performance on the Arena leaderboard.
  • Strong hard-prompt reasoning score on the same leaderboard.
  • Available on five enterprise providers, giving deployment flexibility.

The case against it

  • Premium pricing with no budget tier in its family.
  • Throughput varies sharply by provider: a 64% gap between fastest and slowest measured host.
  • No open weights or fine-tuning access; proprietary licence terms are undisclosed.
00

How good is it?

IntelligencePuzzles, maths, exam questions

4.5 of 5

Arena Text (overall)2nd of 143 · 1497.2

Arena Hard Prompts 2nd of 143Arena Maths 5th of 139LiveBench Reasoning 9th of 35 via Thinking Auto High effortLiveBench Mathematics 16th of 35 via Thinking Auto High effortLiveBench Data Analysis 26th of 35 via Thinking Auto High effort

Also on this board: 1505.4 via Thinking (Aug 2, 2026). Read the pair, not the higher one.

CodingWriting and fixing code on its own

5 of 5

Arena Coding2nd of 143 · 1547.9

Arena Code (WebDev) 11th of 74LiveBench Coding 16th of 35 via Thinking Auto High effort

Also on this board: 1551.1 via Thinking (Aug 2, 2026). Read the pair, not the higher one.

AgenticPlanning, calling tools, staying on task

3.5 of 5

Arena Agent (IPS)8th of 36 · 0.065

SWE-bench Verified 3rd of 39 via mini-SWE-agentLiveBench Agentic Coding 17th of 35 via Thinking Auto High effort

WritingWe do not rate this

Scored, not ratedThe placings are on the right.

Two boards come close and neither tests writing: Arena Creative Writing asks people which of two replies they prefer, and LiveBench Language tests whether a model understood a passage. So we show where Claude Opus 4.6 placed and give it no mark out of five.

Arena Creative Writing 4th of 143 · 1478LiveBench Language 8th of 35 · 83.3 via Thinking Auto High effort
Also scored, on boards we give no mark for
Arena Instruction Following 2nd of 143LiveBench 15th of 35 via Thinking Auto High effortLiveBench Instruction Following 22nd of 35 via Thinking Auto High effort

These tests check whether a model follows instructions — a precondition for all the work above, but not a measure of how well that work is done, which is why they get no rating.

Every published score for this model17 scoresEvery figure we hold, from 17 boards, with who ran it and a link to the source — including the boards no rating above is built on.
LiveBenchreasoning
74.5via Thinking Auto High effortindependentsource ↗
49via Thinking Auto High effortindependentsource ↗
78.2via Thinking Auto High effortindependentsource ↗
69.9via Thinking Auto High effortindependentsource ↗
63.3via Thinking Auto High effortindependentsource ↗
83.3via Thinking Auto High effortindependentsource ↗
89.3via Thinking Auto High effortindependentsource ↗
88.7via Thinking Auto High effortindependentsource ↗
0.065independentsource ↗
1547.9independentsource ↗
1526.7independentsource ↗
1504.5independentsource ↗
1497.2independentsource ↗
1537.9independentsource ↗
75.6via mini-SWE-agentindependentsource ↗
01

Or rent it from someone else

Cheapest published offer

Why this differs from the header. The strip above quotes Anthropic's own list price. This is the cheapest live offer at the widest standard context we hold, whoever is serving it — a reseller undercutting a lab is ordinary commerce, not an error.

per 1M tokens
$5.00 in / $25.00 out
Context served
1M
Throughput
~30 tok/s
Current provider offers with price, context and prompt-privacy answers
ProviderIn / out per 1M tokensContextThroughputTrains on promptsLogs promptsZero retention
Amazon Bedrockus-east-1$5.00 / $25.001M36 tok/sNoNoConfirmed
Microsoft Azure AIglobal$5.00 / $25.001M30 tok/sNoNoUnknown
Microsoft Azure AIus-east-2$5.00 / $25.001M23 tok/sNoNoUnknown
OpenRouter$5.00 / $25.001Mnot measuredUnknownUnknownUnknown
Google Vertex AIglobal$5.00 / $25.001M37 tok/sNoNoConfirmed
Anthropic$5.00 / $25.001M128K out35 tok/sNoYes30 daysUnknown
Google Vertex AIeurope$5.50 / $27.501M36 tok/sNoNoConfirmed

Across the 7 listings we hold: 6 say they do not train on prompts, 0 say they do and 1 do not say. 3 appear in the zero-retention registry we check; the rest are unknown to us rather than confirmed either way.

What each host's API supports

From the parameter list each endpoint publishes. Streaming is omitted: nothing we hold reports it, for any model.

Supported
Not supported
Not published
host gave no parameter list
API features per host
ProviderTool callingJSON outputStrict schema
Amazon Bedrockus-east-1
Microsoft Azure AIglobal
Microsoft Azure AIus-east-2
OpenRouter
Google Vertex AIglobal
Anthropic
Google Vertex AIeurope

Tool calling: 7 of 7 listings say yes. JSON output: 7 of 7 listings say yes. Strict schema: 5 of 7 listings say yes, 2 say no.

02

Models people weigh against Claude Opus 4.6

03

When we formed this view

Dates behind this page

Aug 2, 2026BenchmarkScored 1547.9 on Arena Codingleaderboard
Aug 2, 2026BenchmarkScored 1478 on Arena Creative Writingleaderboard
Aug 2, 2026BenchmarkScored 1526.7 on Arena Hard Promptsleaderboard
Aug 2, 2026BenchmarkScored 1499.4 on Arena Instruction Followingleaderboard
Aug 2, 2026BenchmarkScored 1504.5 on Arena Mathsleaderboard
Aug 2, 2026BenchmarkScored 1497.2 on Arena Text (overall)leaderboard
Aug 2, 2026BenchmarkScored 1537.9 on Arena Code (WebDev)leaderboard
Jul 28, 2026BenchmarkScored 0.065 on Arena Agent (IPS)leaderboard
Jul 26, 2026ListedListed on LLMapfirst indexed by our pipeline
Jun 25, 2026BenchmarkScored 74.5 via Thinking Auto High effort on LiveBenchleaderboard

Prices last checked 9d ago

What we do not know about this model yet

  • Nothing we hold says whether an endpoint streams, so we do not show it either way.
  • 1 of 7 listings do not say whether they train on prompts.
04

Licence and identifiers

What the licence allowsWe hold no licence record for this model. Inside are the identifiers you need to pull it — its Hugging Face repo where we have one, our slug and a machine-readable card.

Licence

Commercial API terms. We hold no licence record for this model, so there is nothing to summarise here.

Identifiers

Modality record
text+image+file->text
Catalogue slug
anthropic-claude-opus-4-6

Machine-readable model card (omc.json) →

Something wrong on this page? Tell us