The AI arena is free today

Open Playground
API Provider32 active models1 organization

OpenAI: API pricing, speed & models

OpenAI hosts 32 active AI models, with input pricing from $0.10 per 1M tokens, with median throughput of 207 characters/sec, and P95 time to first token of 9.19s, with 98.5% success rate over 7 days. Compare OpenAI's API speed, pricing, and reliability against other inference providers.

Image generationVideoFirst-party only
Median throughput207 c/s
Median TTFT2.15s
P95 TTFT9.19s
Success rate (7d)98.5%
From$0.10 /M tok

Catalog

Type
Price
67 models
Model
GPT-5.6 Lunacache pricing
GPT-5.6 Lunacache pricing
GPT-5.6 Terracache pricing
GPT-5.6 Terracache pricing
GPT-5.6 Solcache pricing
GPT-5.6 Solcache pricing
GPT-5.5cache pricing
GPT-5.5cache pricing
GPT-5.450 QPS
GPT-5.450 QPS
GPT-5.2100 QPS
GPT-5.2100 QPS
GPT-5.1100 QPS
GPT-5.1100 QPS
GPT-5 mini200 QPS
GPT-5 mini200 QPS
GPT-4o132 QPS
GPT-4o132 QPS
GPT-4o100 QPS
GPT-4o100 QPS
GPT-4.1100 QPS
GPT-4.1100 QPS
At a glance

OpenAIpricing, performance & catalog

The citable facts about OpenAI's 32 models — sourced from provider APIs and refreshed continuously.

Lowest price
GPT-4.1 nano at $0.100 per 1M input tokens
Highest median throughput
GPT-5.4 mini at 801 chars/s
Lowest median TTFT
GPT-4o at 0.34s
Largest context
GPT-5.6 Luna at 1.1M tokens
Catalog
32 active models from 1 organization

FAQ

Common questions about OpenAI.

What is OpenAI?

OpenAI is an API provider that hosts large language models. Active models: 32; From (input): $0.10 / 1M tok; Median throughput: 207 c/s; Median TTFT: 2.15s; Success rate (7d): 98.5%.

How many models does OpenAI offer?

OpenAI currently serves 32 active models out of 57 historical offerings on LLM Stats.

What is OpenAI's API pricing?

OpenAI input pricing starts from $0.10 per 1M tokens, with the most expensive offering at $10 per 1M tokens. See the Pricing tab above for the full per-model breakdown.

How fast is OpenAI?

OpenAI delivers a median throughput of 207 characters per second and P95 time to first token of 9.19s across its model catalog. See the Models tab for per-model throughput and TTFT breakdowns.

Is OpenAI reliable?

OpenAI has a 98.5% success rate across 5.1K API calls in the last 7 days, with a 1.5% error rate.

Does OpenAI support function calling?

Yes. 44 of 32 models on OpenAI support function calling (tool use). The Capabilities tab lists which specific models accept tool definitions.

Does OpenAI support JSON mode and structured output?

Yes. 44 of 32 models on OpenAI support structured output (JSON mode / schema-constrained generation). The Capabilities tab shows which specific models accept response_format or json_schema parameters.

Does OpenAI offer batch inference?

Yes. 44 of 32 models on OpenAI support batch inference for cheaper, asynchronous workloads.

Does OpenAI support fine-tuning?

Yes. 10 of 32 models on OpenAI support fine-tuning. The Capabilities tab lists which specific models accept fine-tuning jobs.

Does OpenAI support multimodal models?

Yes. OpenAI's catalog includes 31 vision-capable, 12 image generation, 2 audio, and 4 video models. See the Models and Capabilities tabs for the full per-model breakdown.

Whose models does OpenAI host?

OpenAI hosts models from OpenAI. See the Models tab for the full catalog grouped by creator.

How do I start using OpenAI?

Sign up at https://openai.com to get an API key, then call OpenAI's API directly from your application. Most clients work out of the box by pointing the OpenAI SDK at OpenAI's base URL with your key. Use the Models and Pricing tabs above to pick the right model for your latency, cost, and context-window requirements.