The AI arena is free today

Open Superagent
AnthropicReleased on Aug 5, 2025

Claude Opus 4.1: API Pricing, Context Window & Benchmarks

Claude Opus 4.1 is a language model from Anthropic, released in August 2025, with multimodal input.

Claude Opus 4.1 is a hybrid reasoning model that pushes the frontier for coding and AI agents, featuring a 200K context window. It delivers superior performance and precision for real-world coding and agentic tasks, handling complex

Input
TextImage
Output
Text

Claude Opus 4.1 benchmarks

Capability tiers

Standing within each category, adjusted for leaderboard depth.

Real tasks performance

How Claude Opus 4.1 performs across real-world prompt categories.

Performance by conversation depth

How Claude Opus 4.1 holds up as conversations get longer.

Quality Tracker

Claude Opus 4.1 Performance Across Datasets

Scores sourced from the model's scorecard, paper, or official blog posts

LLM Stats Logollm-stats.com - Mon Aug 24 2026
Notice missing or incorrect data?

Claude Opus 4.1 pricing

Providers

Claude Opus 4.1 starts at $15.00 per million input tokens and $75.00 per million output tokens via Anthropic. See all 3 providers below with their per-token pricing, latency, throughput, and modality support.

ProviderInput $/MCached input $/MOutput $/MContext in / outTTFT p95 sOutput p5 c/sModalities in / out
Anthropic logoAnthropic
$15.00$75.00200.0K/32.0K
0.50
/
Bedrock logoBedrock
$15.00$75.00200.0K/32.0K
0.50
/
Google logoGoogle
$15.00$75.00200.0K/32.0K
0.40
/

Cached input is the discounted price for prompt tokens served from a provider cache. TTFT is time to first token. Output is characters per second; p5 is the sustained floor exceeded by 95% of observed requests.

Loading chart...
Loading chart...
Loading chart...

Claude Opus 4.1 context window

Input and output token limits for Claude Opus 4.1, plus how it ranks on long-context understanding.

InputOutput
200Ktokens
32Ktokens
301 pages of text
200K
8K128K1M

Claude Opus 4.1 API

Available from the model provider

Claude Opus 4.1 has an official provider API. It is not currently routed through the LLM Stats gateway.

Read the official API documentation

Claude Opus 4.1 latency

Claude Opus 4.1 time to first token, sustained output throughput, and failed-request rate from live API traffic over the trailing 7 days.

Claude Opus 4.1 examples

Recent arena outputs from Claude Opus 4.1, picked from the highest-ranked matchups.

Claude Opus 4.1 license

Claude Opus 4.1 is a proprietary model available under its provider's product and API terms.

License
Proprietary
Hosted access

Proprietary license - usage restrictions apply

Claude Opus 4.1 resources

Official sources for Claude Opus 4.1: api documentation, official playground, official launch post.

Claude Opus 4.1 vs other models

The most-compared alternatives to Claude Opus 4.1 are GPT OSS 120B High, Phi 4 Reasoning Plus, Step-3.5-Flash. Open any pair side-by-side for benchmarks, pricing, context, and latency.

Models like Claude Opus 4.1

Models ranked just above and below Claude Opus 4.1 by LLM Stats score.

 

GPT OSS 120B High

Score pending
 

Phi 4 Reasoning Plus

Score pending
 

Step-3.5-Flash

Score pending
 

GPT-5.2

Score pending
 

Claude Sonnet 4.5

Score pending
 

LongCat-Flash-Thinking-2601

Score pending

FAQ

Common questions about Claude Opus 4.1.

When was Claude Opus 4.1 released?

Claude Opus 4.1 was released on August 5, 2025 by Anthropic. This is the official Claude Opus 4.1 release date tracked on LLM Stats.

How much does Claude Opus 4.1 cost?

Claude Opus 4.1 pricing starts at $15.00 per million input tokens and $75.00 per million output tokens via Anthropic, the lowest price among tracked providers.

Is Claude Opus 4.1 available via API?

Yes, Claude Opus 4.1 is available via API. See the official documentation for authentication and endpoint details. It is served by 3 providers tracked on LLM Stats.

Who created Claude Opus 4.1?

Claude Opus 4.1 was created by Anthropic.

What is the license for Claude Opus 4.1?

Claude Opus 4.1 is released under the Proprietary license.

Is Claude Opus 4.1 multimodal?

Yes, Claude Opus 4.1 is multimodal and can accept both text and images as input.

What is Claude Opus 4.1 latency?

Claude Opus 4.1 p95 time to first token is 0.40 seconds via Google over the trailing 7 days. Lower time to first token means the model begins responding sooner for chat, agents and API workloads.

Where can I use Claude Opus 4.1?

Claude Opus 4.1 is available through 3 providers including Anthropic, Bedrock, Google.

Where is the Claude Opus 4.1 paper or technical report?

Claude Opus 4.1 has a paper or technical report available at https://www.anthropic.com/news/claude-opus-4-1. Use that source for architecture, training, release and evaluation details.

What models should I compare Claude Opus 4.1 against?

Common Claude Opus 4.1 comparisons include Claude Opus 4.1 vs GPT OSS 120B High, Claude Opus 4.1 vs Phi 4 Reasoning Plus, Claude Opus 4.1 vs Step-3.5-Flash. Compare them side by side for benchmark scores, pricing, context window, latency and API availability.