The AI arena is free today

Open Superagent
API Provider10 active models1 organization

Anthropic: API pricing, speed & models

Anthropic hosts 10 active AI models, with input pricing from $1.00 per 1M tokens, with median throughput of 46 characters/sec, and P95 time to first token of 8.24s, with 99.4% success rate over 7 days. Compare Anthropic's API speed, pricing, and reliability against other inference providers.

First-party only
Median throughput46 c/s
Median TTFT2.23s
P95 TTFT8.24s
Success rate (7d)99.4%
From$1.00 /M tok

Catalog

Type
Price
24 models
Model
Claude Fable 5.1cache pricing
Claude Fable 5.1cache pricing
Claude Opus 5cache pricing
Claude Opus 5cache pricing
Claude Sonnet 5cache pricing · 42 QPS
Claude Sonnet 5cache pricing · 42 QPS
Claude Fable 5cache pricing
Claude Fable 5cache pricing
Claude Opus 4.8cache pricing · 42 QPS
Claude Opus 4.8cache pricing · 42 QPS
At a glance

Anthropicpricing, performance & catalog

The citable facts about Anthropic's 10 models — sourced from provider APIs and refreshed continuously.

Lowest price
Claude Haiku 4.5 at $1.00 per 1M input tokens
Highest median throughput
Claude Opus 5 at 115 chars/s
Lowest median TTFT
Claude Haiku 4.5 at 0.69s
Largest context
Claude Fable 5.1 at 1.0M tokens
Catalog
10 active models from 1 organization

FAQ

Common questions about Anthropic.

What is Anthropic?

Anthropic is an API provider that hosts large language models. Active models: 10; From (input): $1.00 / 1M tok; Median throughput: 46 c/s; Median TTFT: 2.23s; Success rate (7d): 99.4%.

How many models does Anthropic offer?

Anthropic currently serves 10 active models out of 22 historical offerings on LLM Stats.

What is Anthropic's API pricing?

Anthropic input pricing starts from $1.00 per 1M tokens, with the most expensive offering at $10 per 1M tokens. See the Pricing tab above for the full per-model breakdown.

How fast is Anthropic?

Anthropic delivers a median throughput of 46 characters per second and P95 time to first token of 8.24s across its model catalog. See the Models tab for per-model throughput and TTFT breakdowns.

Is Anthropic reliable?

Anthropic has a 99.4% success rate across 1.9K API calls in the last 7 days, with a 0.6% error rate.

Does Anthropic support function calling?

Yes. 24 of 10 models on Anthropic support function calling (tool use). The Capabilities tab lists which specific models accept tool definitions.

Does Anthropic support JSON mode and structured output?

Yes. 24 of 10 models on Anthropic support structured output (JSON mode / schema-constrained generation). The Capabilities tab shows which specific models accept response_format or json_schema parameters.

Does Anthropic offer batch inference?

Yes. 24 of 10 models on Anthropic support batch inference for cheaper, asynchronous workloads.

Does Anthropic support multimodal models?

Yes. Anthropic's catalog includes 10 vision-capable models. See the Models and Capabilities tabs for the full per-model breakdown.

Whose models does Anthropic host?

Anthropic hosts models from Anthropic. See the Models tab for the full catalog grouped by creator.

How do I start using Anthropic?

Sign up at https://anthropic.com to get an API key, then call Anthropic's API directly from your application. Most clients work out of the box by pointing the OpenAI SDK at Anthropic's base URL with your key. Use the Models and Pricing tabs above to pick the right model for your latency, cost, and context-window requirements.