The AI arena is free today

Open Superagent
OpenAIReleased on Sep 12, 2024

o1-mini: Benchmarks, Pricing & Context Window

o1-mini is a language model from OpenAI, released in September 2024.

o1-mini is a cost-efficient language model developed by OpenAI, designed for advanced reasoning tasks while minimizing computational resources.

Input
Text
Output
Text

o1-mini benchmarks

Capability tiers

Standing within each category, adjusted for leaderboard depth.

Real tasks performance

High-confidence performance for o1-mini across real-world prompt categories. Only 95% intervals at most 4 points wide are shown.

Performance by conversation depth

How o1-mini holds up as conversations get longer.

Quality Tracker

o1-mini Performance Across Datasets

Scores sourced from the model's scorecard, paper, or official blog posts

LLM Stats Logollm-stats.com - Tue Sep 15 2026
Notice missing or incorrect data?

o1-mini pricing

Providers

o1-mini starts at $3.00 per million input tokens and $12.00 per million output tokens via OpenAI. See all 2 providers below with their per-token pricing, latency, throughput, and modality support.

ProviderInput $/MCached input $/MOutput $/MContext in / outTTFT p95 sOutput p5 c/sModalities in / out
OpenAI logoOpenAI
$3.00$12.00128.0K/65.5K
5.20
/
Azure logoAzure
$3.30$13.20128.0K/65.5K
0.50
/

Cached input is the discounted price for prompt tokens served from a provider cache. TTFT is time to first token. Output is characters per second; p5 is the sustained floor exceeded by 95% of observed requests.

Loading chart...
Loading chart...
Loading chart...

o1-mini context window

Input and output token limits for o1-mini, plus how it ranks on long-context understanding.

InputOutput
128Ktokens
66Ktokens
192 pages of text
128K
8K128K1M

Try now

huggle
o1-miniin Huggle

Make it with
o1-mini.

o1-mini

o1-mini latency

o1-mini time to first token, sustained output throughput, and failed-request rate from live model usage over the trailing 7 days.

o1-mini examples

Recent arena outputs from o1-mini, picked from the highest-ranked matchups.

o1-mini license

o1-mini is a proprietary model available under its provider's product and API terms.

License
Proprietary
Hosted access

Proprietary license - usage restrictions apply

o1-mini resources

Official sources for o1-mini: provider documentation, official playground, paper or system card, official launch post.

o1-mini vs other models

The most-compared alternatives to o1-mini are Claude 3.5 Sonnet, DeepSeek R1 Distill Qwen 14B, Qwen3 Max. Open any pair side-by-side for benchmarks, pricing, context, and latency.

Models like o1-mini

Models ranked just above and below o1-mini by LLM Stats score.

 

Claude 3.5 Sonnet

Score pending
 

DeepSeek R1 Distill Qwen 14B

Score pending
 

Qwen3 Max

Score pending
 

Qwen3 VL 8B Thinking

Score pending
 

DeepSeek-V3

Score pending
 

Kimi K2 Instruct

Score pending

FAQ

Common questions about o1-mini.

When was o1-mini released?

o1-mini was released on September 12, 2024 by OpenAI. This is the official o1-mini release date tracked on LLM Stats.

How much does o1-mini cost?

o1-mini pricing starts at $3.00 per million input tokens and $12.00 per million output tokens via OpenAI, the lowest price among tracked providers.

Who created o1-mini?

o1-mini was created by OpenAI.

What is the license for o1-mini?

o1-mini is released under the Proprietary license.

What is o1-mini latency?

o1-mini p95 time to first token is 0.50 seconds via Azure over the trailing 7 days. Lower time to first token means the model begins responding sooner for chat, agents and model workloads.

Where can I use o1-mini?

o1-mini is available through 2 providers including OpenAI, Azure.

Where is the o1-mini paper or technical report?

o1-mini has a paper or technical report available at https://cdn.openai.com/o1-system-card-20240917.pdf. Use that source for architecture, training, release and evaluation details.

What models should I compare o1-mini against?

Common o1-mini comparisons include o1-mini vs Claude 3.5 Sonnet, o1-mini vs DeepSeek R1 Distill Qwen 14B, o1-mini vs Qwen3 Max. Compare them side by side for benchmark scores, pricing, context window, latency and provider availability.