Models / Anthropic/ Claude Opus 4.8

Claude Opus 4.8

Anthropic · released May 27, 2026

Input: text, images and documents. Output: text.InputOutput
Type
Closed
Input
$5.00
Output
$25.00
Cached
$0.50

List price · per 1M tokens · Anthropic at 1M context · machine-readable source ↗

Our take

Written Sep 30, 2026

Claude Opus 4.8 is a hosted-only model with no download we can point you to, so using it means choosing a host. Its measured strengths sit in coding and instruction following, and its weakest showing is tool use inside agent sessions.

Who should pick it

Reach for it when the work is coding or tightly specified instruction following, where its measured standings are highest, or when long documents would otherwise need splitting up first. It is a hosted model, so you reach it through a host rather than running it yourself. Skip it if you need reliable tool use inside an agent session, or if you need licence terms you can read before you build on it.

The case for it

  • 14th of 168 on Arena Coding as of 25 Sep 2026, with 11th of 58 on LiveBench Coding via Max effort as of 25 Jun 2026 — code generation and completion rather than fixing issues in an existing project.
  • 15th of 168 on Arena Instruction Following as of 25 Sep 2026, so constrained-rewriting work is a reasonable place to start.
  • 3rd of 55 on Arena Agent · Recovery as of 25 Sep 2026, an inverse-propensity score for getting back on track after a command fails.
  • Request capacity leaves room for long documents without splitting them up first, though reliable recall across all of it is unverified in our data.

The case against it

  • 54th of 55 on Arena Agent · Tool use as of 25 Sep 2026, so do not put it at the centre of an agent session that depends on calling the right tool.
  • 50th of 58 on LiveBench Data Analysis via Max effort as of 25 Jun 2026, which covers table and event-ordering tasks.
  • We list no download for it and no licence, so permissions for commercial use, changes and redistribution are unverified.
00

How good is it?

A closed text model for everyday questions, writing and coding, though calling tools is where it struggles.

Good at
  • getting answers to everyday questionsArena Text (overall) · 27th of 168
  • drafts, rewrites and editingArena Creative Writing · 16th of 168
  • writing and completing codeArena Coding · 14th of 168
  • changing course when you give new instructionsArena Agent · Steerability · 4th of 55
  • getting back on track after a step failsArena Agent · Recovery · 3rd of 55
Less good at
  • calling tools to carry out requestsArena Agent · Tool use · 54th of 55

EverydayGeneral questions and everyday reasoning

4 of 5

Arena Text (overall)27th of 168 · 1474

Arena Hard Prompts 16th of 168Arena Maths 28th of 163LiveBench Mathematics 14th of 58LiveBench Reasoning 15th of 58LiveBench Data Analysis 50th of 58

Also on this board: 1480 (Sep 25, 2026), 1482 (Aug 10, 2026). Read the pair, not the higher one.

CodingWriting and fixing code on its own

4.5 of 5

Arena Coding14th of 168 · 1530

Arena Code (WebDev) 31st of 95LiveBench Coding 11th of 58

Also on this board: 1534 (Sep 25, 2026), 1533 (Aug 10, 2026). Read the pair, not the higher one.

AgenticPlanning, calling tools, staying on task

2.5 of 5

Arena Agent19th of 55 · 0.019

LiveBench Agentic Coding 35th of 58

Also on this board: 0.073 (Sep 25, 2026), 0.095 (Aug 11, 2026). Read the pair, not the higher one.

WritingDrafting and rewriting prose

3.5 of 5

Arena Creative Writing16th of 168 · 1462

LiveBench Language 31st of 58

Also on this board: 1467 (Sep 25, 2026), 1467 (Aug 10, 2026). Read the pair, not the higher one.

How it behaves in an agent loop
Tool usereaches for the right one, and does not invent one54th of 55
Steerabilitydoes what it was asked, and changes course when told4th of 55
Recoverygets back on track after a command fails3rd of 55
Task outcomefinishes what the session set out to do14th of 55

Placings on Arena's agent boards, from live sessions people ran themselves. A model can lead on one of these and sit mid-field on the others.

Other boards it appears on
Arena Instruction Following 15th of 168LiveBench Instruction Following 17th of 58LiveBench 23rd of 58Arena Agent · Recovery 3rd of 55Arena Agent · Steerability 4th of 55Arena Agent · Task outcome 14th of 55Arena Agent · Tool use 54th of 55

Boards this model appears on that none of the ratings above are built on.

Every published score for this model20 scoresEvery figure we hold, from 20 boards, with who ran it and a link to the source — including the boards no rating above is built on.
LiveBenchreasoning
76.22source ↗
50.51source ↗
81.83source ↗
66.03source ↗
72.03source ↗
79.66source ↗
94.32source ↗
89.19source ↗
0.019source ↗
0.1source ↗
0.094source ↗
0.059source ↗
−0.272source ↗
1530source ↗
1462source ↗
1504source ↗
1477source ↗
1474source ↗
1535source ↗
01

Where to rent it

Prices checked 3 hours ago — each listing carries its own date.

Cheapest published offer

Anthropic, direct

The lab is also the cheapest we hold. The strip above and this offer are the same one, so nothing on this page undercuts Anthropic on 1M of context.

per 1M tokens
$5.00 in / $25.00 out
Context served
1M
Throughput
~54 tok/s
Current provider offers with price, context and prompt-privacy answers
ProviderIn / out per 1M tokensContextThroughputTrains on promptsLogs promptsZero retention
OpenRouterOpenRouter's own listing$5.00 / $25.00checked 3 hours ago1Mnot measuredUnknownUnknownUnknown
DeepInfraDirect$5.00 / $25.00checked 3 hours ago1Mnot measuredUnknownUnknownUnknown
Microsoft Azure AIglobalThrough OpenRouter$5.00 / $25.00checked 3 hours ago1M128K max reply23 tok/sNoNoUnknown
Google Vertex AIglobalThrough OpenRouter$5.00 / $25.00checked 3 hours ago1M128K max reply75 tok/sNoNoConfirmed
AnthropicDirect$5.00 / $25.00checked 3 hours ago1M128K max reply54 tok/sNoYes30 daysUnknown
Amazon BedrockThrough OpenRouter$5.00 / $25.00checked 3 hours ago1M128K max reply58 tok/sNoNoConfirmed
Claude Platform on AWSThrough OpenRouter$5.00 / $25.00checked 3 hours ago1M128K max reply41 tok/sNoYes30 daysUnknown
Amazon BedrockusThrough OpenRouter$5.50 / $27.50checked 3 hours ago1M128K max reply65 tok/sNoNoConfirmed
Microsoft Azure AIusThrough OpenRouter$5.50 / $27.50checked 3 hours ago1M128K max reply29 tok/sNoNoUnknown
Amazon Bedrockeu-west-1Through OpenRouter$5.50 / $27.50checked 3 hours ago1M128K max reply73 tok/sNoNoConfirmed
Google Vertex AIeuropeThrough OpenRouter$5.50 / $27.50checked 3 hours ago1M128K max reply110 tok/sNoNoConfirmed
Google Vertex AIusThrough OpenRouter$5.50 / $27.50checked 3 hours ago1M128K max reply84 tok/sNoNoConfirmed
Anthropicfast tierThrough OpenRouter$10.00 / $50.00checked 3 hours ago1M128K max reply133 tok/sNoYes30 daysUnknown

Across the 13 listings we hold: 11 say they do not train on prompts, 0 say they do and 2 do not say. 6 appear in the zero-retention registry we check; the rest are unknown to us.

What each host's API supports

From the parameter list each endpoint publishes. Streaming is omitted: nothing we hold reports it, for any model.

API features per host
ProviderTool callingJSON outputStrict schema
OpenRouterOpenRouter's own listing✓✓✓
DeepInfraDirect
Microsoft Azure AIglobalThrough OpenRouter✓✓✓
Google Vertex AIglobalThrough OpenRouter✓✓✓
AnthropicDirect✓✓✓
Amazon BedrockThrough OpenRouter✓✓✗
Claude Platform on AWSThrough OpenRouter✓✓✓
Amazon BedrockusThrough OpenRouter✓✓✗
Microsoft Azure AIusThrough OpenRouter✓✓✓
Amazon Bedrockeu-west-1Through OpenRouter✓✓✗
Google Vertex AIeuropeThrough OpenRouter✓✓✓
Google Vertex AIusThrough OpenRouter✓✓✓
AnthropicfastThrough OpenRouter✓✓✓

Tool calling: 12 of 13 listings say yes, 1 publishes no parameter list. JSON output: 12 of 13 listings say yes, 1 publishes no parameter list. Strict schema: 9 of 13 listings say yes, 3 say no, 1 publishes no parameter list.

02

Models people weigh against Claude Opus 4.8

03

When we formed this view

Recent changes

Sep 25, 2026BenchmarkScored 1530 on Arena Coding
What movedleaderboard
Sep 25, 2026BenchmarkScored 1462 on Arena Creative Writing
What movedleaderboard
Sep 25, 2026BenchmarkScored 1504 on Arena Hard Prompts
What movedleaderboard
Sep 25, 2026BenchmarkScored 1478 on Arena Instruction Following
What movedleaderboard
Sep 25, 2026BenchmarkScored 1477 on Arena Maths
What movedleaderboard
Sep 25, 2026BenchmarkScored 1474 on Arena Text (overall)
What movedleaderboard
Sep 25, 2026BenchmarkScored 1535 on Arena Code (WebDev)
What movedleaderboard
Sep 21, 2026Price changeHost Azure (US region) raised Claude Opus 4.8 pricing by 10% on all rates
What movedinput +10% ($5.00 → $5.50 per 1M tokens), output +10% ($25.00 → $27.50 per 1M tokens), cache read +10% ($0.50 → $0.55 per 1M tokens), cache write +10% ($6.25 → $6.88 per 1M tokens)
Sep 5, 2026BenchmarkScored 0.019 on Arena Agent
What movedleaderboard
Sep 5, 2026BenchmarkScored 0.1 on Arena Agent · Recovery
What movedleaderboard

Each date is the day we first saw the change, or the day the maker announced it.

What we do not know about this model yet

  • 1 of 13 listings publishes no parameter list, so what its API accepts is unknown to us.
  • Nothing we hold says whether an endpoint streams, so we do not show it either way.
  • 2 of 13 listings do not say whether they train on prompts.
  • We hold no batch or off-peak rate for any of its listings.
04

Licence and identifiers

What the licence allowsWe hold no licence record for this model. Inside are the identifiers you need to pull it — its Hugging Face repo where we have one, our slug and a machine-readable card.

Licence

We hold no licence record for this model, and no record of published weights either — so we can neither summarise its terms nor point you at the weights.

Identifiers

Takes in, gives back
Text, images and documents in, text out
Catalogue slug
anthropic-claude-opus-4-8

Machine-readable model card (omc.json) →

Something wrong on this page? Tell us