OpenAIReleased on Dec 17, 2024

o1: API Pricing, Context Window & Benchmarks

o1 is a language model from OpenAI, released in December 2024.

A research preview model focused on mathematical and logical reasoning capabilities, demonstrating improved performance on tasks requiring step-by-step reasoning, mathematical problem-solving, and code generation. The model shows enhanced

Input
Text
Output
Text

o1 benchmarks

Rankings

Quality Tracker

o1 Performance Across Datasets

Scores sourced from the model's scorecard, paper, or official blog posts

LLM Stats Logollm-stats.com - Tue Jul 28 2026
Notice missing or incorrect data?

o1 pricing

Providers

o1 starts at $15.00 per million input tokens and $60.00 per million output tokens via Azure. See all 2 providers below with their per-token pricing, latency, throughput, and modality support.

ProviderInput $/MOutput $/MContext in / outTTFT p50 / p95 sOutput avg / p5 c/sSuccess 7dModalities in / out
Azure logoAzure
$15.00$60.00200.0K/100.0K
/0.54
16/
/
OpenAI logoOpenAI
$15.00$60.00200.0K/100.0K
/16.20
66/
/

Cached input is the discounted price for prompt tokens served from a provider cache. TTFT is time to first token. Output is characters per second; p5 is the sustained floor exceeded by 95% of observed requests. Success is calculated from completed versus failed requests over the trailing seven days.

Loading chart...
Loading chart...
Loading chart...

o1 context window

Input and output token limits for o1, plus how it ranks on long-context understanding.

InputOutput
200Ktokens
100Ktokens
301 pages of text
200K
8K128K1M

o1 API

Available from the model provider

o1 has an official provider API. It is not currently routed through the LLM Stats gateway.

Read the official API documentation

o1 latency

o1 time to first token, sustained output throughput, and failed-request rate from live API traffic over the trailing 7 days.

o1 examples

Recent arena outputs from o1, picked from the highest-ranked matchups.

o1 license

o1 is a proprietary model available under its provider's product and API terms.

License
Proprietary
Hosted access

Proprietary license - usage restrictions apply

o1 resources

Official sources for o1: api documentation, paper or system card, official launch post, source repository.

o1 vs other models

The most-compared alternatives to o1 are Mistral Large 3, Sarvam-105B, Qwen3-235B-A22B-Instruct-2507. Open any pair side-by-side for benchmarks, pricing, context, and latency.

Models like o1

Models ranked just above and below o1 by LLM Stats score.

 

Mistral Large 3

Score pending
 

Sarvam-105B

Score pending
 

Qwen3-235B-A22B-Instruct-2507

Score pending
 

GPT-5

Score pending
 

DeepSeek-V3

Score pending
 

o1-preview

Score pending

FAQ

Common questions about o1.

When was o1 released?

o1 was released on December 17, 2024 by OpenAI. This is the official o1 release date tracked on LLM Stats.

How much does o1 cost?

o1 pricing starts at $15.00 per million input tokens and $60.00 per million output tokens via Azure, the lowest price among tracked providers.

Is o1 available via API?

Yes, o1 is available via API. See the official documentation for authentication and endpoint details. It is served by 2 providers tracked on LLM Stats.

Who created o1?

o1 was created by OpenAI.

What is the license for o1?

o1 is released under the Proprietary license.

What is o1 latency?

o1 p95 time to first token is 0.54 seconds via Azure over the trailing 7 days. Lower time to first token means the model begins responding sooner for chat, agents and API workloads.

Where can I use o1?

o1 is available through 2 providers including Azure, OpenAI.

Where is the o1 paper or technical report?

o1 has a paper or technical report available at https://cdn.openai.com/o1-system-card-20240917.pdf. Use that source for architecture, training, release and evaluation details.

What models should I compare o1 against?

Common o1 comparisons include o1 vs Mistral Large 3, o1 vs Sarvam-105B, o1 vs Qwen3-235B-A22B-Instruct-2507. Compare them side by side for benchmark scores, pricing, context window, latency and API availability.