The AI arena is free today

Open Superagent
AnthropicReleased on May 28, 2026

Claude Opus 4.8: Benchmarks, Pricing & Context Window

Claude Opus 4.8 is a language model from Anthropic, released in May 2026, with multimodal input, a 1M-token context window, and pricing from $5.00/M input, $0.500/M cached input, $25.00/M output.

Claude Opus 4.8 is Anthropic's upgrade to Opus 4.7 and its most capable general-access model at release, with improvements across software engineering, agentic tool use, reasoning, computer use, and knowledge-work benchmarks while shipping

Input
TextImage
Output
Text

Claude Opus 4.8 benchmarks

Capability tiers

Standing within each category, adjusted for leaderboard depth.

Real tasks performance

High-confidence performance for Claude Opus 4.8 across real-world prompt categories. Only 95% intervals at most 4 points wide are shown.

Performance by conversation depth

How Claude Opus 4.8 holds up as conversations get longer.

Quality Tracker

Claude Opus 4.8 Performance Across Datasets

Scores sourced from the model's scorecard, paper, or official blog posts

LLM Stats Logollm-stats.com - Sun Oct 11 2026
Notice missing or incorrect data?

Claude Opus 4.8 pricing

Providers

Claude Opus 4.8 starts at $5.00 per million input tokens and $25.00 per million output tokens via Anthropic. Reused prompt prefixes cost $0.500 per million cached input tokens. See all 3 providers below with their per-token pricing, latency, throughput, and modality support.

ProviderInput $/MCached input $/MOutput $/MContext in / outTTFT p95 sOutput p5 c/sModalities in / out
Anthropic logoAnthropic
$5.00$0.500$25.001.0M/128.0K
0.50
—
/
DeepInfra logoDeepInfra
$5.00—$25.001.0M/1.0M
—
—
/
Vertex AI logoVertex AI
$5.00—$25.001.0M/128.0K
0.50
—
/

Cached input is the discounted price for prompt tokens served from a provider cache. TTFT is time to first token. Output is characters per second; p5 is the sustained floor exceeded by 95% of observed requests.

Loading chart...
Loading chart...
Loading chart...
Loading chart...

Claude Opus 4.8 context window

Input and output token limits for Claude Opus 4.8, plus how it ranks on long-context understanding.

InputOutput
1Mtokens
1Mtokens
≈ 1.5k pages of text
1M
8K128K1M

Try now

huggle
Claude Opus 4.8in Huggle

Make it with
Claude Opus 4.8.

Claude Opus 4.8

Claude Opus 4.8 latency

Claude Opus 4.8 time to first token, sustained output throughput, and failed-request rate from live model usage over the trailing 7 days.

Claude Opus 4.8 examples

Recent arena outputs from Claude Opus 4.8, picked from the highest-ranked matchups.

Claude Opus 4.8 license

Claude Opus 4.8 is a proprietary model available under its provider's product and API terms.

License
Proprietary
Hosted access

Proprietary license - usage restrictions apply

Claude Opus 4.8 resources

Official sources for Claude Opus 4.8: provider documentation, official playground, paper or system card, official launch post.

Claude Opus 4.8 vs other models

The most-compared alternatives to Claude Opus 4.8 are Claude Opus 4.6, Claude Mythos Preview, Ember-1. Open any pair side-by-side for benchmarks, pricing, context, and latency.

Models like Claude Opus 4.8

Models ranked just above and below Claude Opus 4.8 by LLM Stats score.

 

Claude Opus 4.6

Score pending
 

Claude Mythos Preview

Score pending
 

Ember-1

Score pending
 

Claude Opus 4.7

Score pending
 

Kimi K3

Score pending
 

GPT-5.5

Score pending

FAQ

Common questions about Claude Opus 4.8.

When was Claude Opus 4.8 released?

Claude Opus 4.8 was released on May 28, 2026 by Anthropic. This is the official Claude Opus 4.8 release date tracked on LLM Stats.

How much does Claude Opus 4.8 cost?

Claude Opus 4.8 pricing starts at $5.00 per million input tokens, $0.50 per million cached input tokens, $25.00 per million output tokens via Anthropic, the lowest price among tracked providers.

Who created Claude Opus 4.8?

Claude Opus 4.8 was created by Anthropic.

What is the license for Claude Opus 4.8?

Claude Opus 4.8 is released under the Proprietary license.

Is Claude Opus 4.8 multimodal?

Yes, Claude Opus 4.8 is multimodal and can accept both text and images as input.

What is Claude Opus 4.8 latency?

Claude Opus 4.8 p95 time to first token is 0.50 seconds via Anthropic over the trailing 7 days. Lower time to first token means the model begins responding sooner for chat, agents and model workloads.

Where can I use Claude Opus 4.8?

Claude Opus 4.8 is available through 3 providers including Anthropic, DeepInfra, Vertex AI.

Where is the Claude Opus 4.8 paper or technical report?

Claude Opus 4.8 has a paper or technical report available at https://cdn.sanity.io/files/4zrzovbb/website/c886650a2e96fc0925c805a1a7ca77314ccbf4a6.pdf. Use that source for architecture, training, release and evaluation details.

What models should I compare Claude Opus 4.8 against?

Common Claude Opus 4.8 comparisons include Claude Opus 4.8 vs Claude Opus 4.6, Claude Opus 4.8 vs Claude Mythos Preview, Claude Opus 4.8 vs Ember-1. Compare them side by side for benchmark scores, pricing, context window, latency and provider availability.