o4-mini: API Pricing, Context Window & Benchmarks
o4-mini is a language model from OpenAI, released in April 2025, with multimodal input.
o4-mini is OpenAI's latest small o-series model, optimized for fast, effective reasoning with exceptionally efficient performance in coding and visual tasks. It is faster and more affordable than o3.
o4-mini benchmarks
Rankings
Quality Tracker
o4-mini Performance Across Datasets
Scores sourced from the model's scorecard, paper, or official blog posts
o4-mini pricing
Providers
o4-mini starts at $1.10 per million input tokens and $4.40 per million output tokens via OpenAI.
| Provider | Input $/M | Output $/M | Context in / out | TTFT p50 / p95 s | Output avg / p5 c/s | Success 7d | Modalities in / out |
|---|---|---|---|---|---|---|---|
| $1.10 | $4.40 | 200.0K/100.0K | —/5.20 | 115/— | — | / |
Cached input is the discounted price for prompt tokens served from a provider cache. TTFT is time to first token. Output is characters per second; p5 is the sustained floor exceeded by 95% of observed requests. Success is calculated from completed versus failed requests over the trailing seven days.
o4-mini context window
Input and output token limits for o4-mini, plus how it ranks on long-context understanding.
o4-mini API
Available from the model provider
o4-mini has an official provider API. It is not currently routed through the LLM Stats gateway.
Read the official API documentationo4-mini latency
o4-mini time to first token, sustained output throughput, and failed-request rate from live API traffic over the trailing 7 days.
o4-mini examples
Recent arena outputs from o4-mini, picked from the highest-ranked matchups.
o4-mini license
o4-mini is a proprietary model available under its provider's product and API terms, has a knowledge cutoff of May 2024.
- License
- Proprietary
- Hosted access
- Knowledge cutoff
- May 2024
Proprietary license - usage restrictions apply
o4-mini resources
Official sources for o4-mini: api documentation, paper or system card, source repository.
o4-mini vs other models
The most-compared alternatives to o4-mini are Seed 2.0 Lite, K-EXAONE-236B-A23B, LongCat-Flash-Thinking. Open any pair side-by-side for benchmarks, pricing, context, and latency.
Models like o4-mini
Models ranked just above and below o4-mini by LLM Stats score.
FAQ
Common questions about o4-mini.