The AI arena is free today

Open Playground
OpenAIReleased on May 5, 2026

GPT-5.5 Instant: API Pricing, Context Window & Benchmarks

GPT-5.5 Instant is a language model from OpenAI, released in May 2026, with multimodal input, a 400K-token context window, and pricing from $5.00/M input and $30.00/M output.

GPT-5.5 Instant is OpenAI's latest Instant model and the default ChatGPT model released on May 5, 2026. It improves factuality, everyday reasoning, image understanding, STEM answers, concision, conversational tone, and personalization over

Input
TextImage
Output
Text

GPT-5.5 Instant benchmarks

Rankings

Quality Tracker

GPT-5.5 Instant Performance Across Datasets

Scores sourced from the model's scorecard, paper, or official blog posts

LLM Stats Logollm-stats.com - Thu Aug 06 2026
Notice missing or incorrect data?

GPT-5.5 Instant pricing

Providers

GPT-5.5 Instant starts at $5.00 per million input tokens and $30.00 per million output tokens via OpenAI.

ProviderInput $/MCached input $/MOutput $/MContext in / outTTFT p95 sOutput p5 c/sModalities in / out
OpenAI logoOpenAI
$5.00$30.00400.0K/128.0K
1.33
32
/

Cached input is the discounted price for prompt tokens served from a provider cache. TTFT is time to first token. Output is characters per second; p5 is the sustained floor exceeded by 95% of observed requests.

GPT-5.5 Instant context window

Input and output token limits for GPT-5.5 Instant, plus how it ranks on long-context understanding.

InputOutput
400Ktokens
128Ktokens
602 pages of text
400K
8K128K1M

GPT-5.5 Instant API

POST/v1/chat/completions

Run a request to see the response

Use it in your code

Billed at $5.00 input / $30.00 output per 1M tokens through the LLM Stats gateway.

from openai import OpenAI

client = OpenAI(
    api_key="YOUR_API_KEY",
    base_url="https://gateway.llm-stats.com/v1"
)

response = client.chat.completions.create(
    model="gpt-5.5-instant",
    messages=[
        {"role": "user", "content": "What is machine learning?"}
    ]
)

print(response.choices[0].message.content)

Need an API key? Create one above in the playground, or read the API documentation.

GPT-5.5 Instant latency

GPT-5.5 Instant time to first token, sustained output throughput, and failed-request rate from live API traffic over the trailing 7 days.

Provider operational metrics

Time to first token, output throughput, and failed-request rate from live API traffic

Loading chart...
Loading chart...
Loading chart...

GPT-5.5 Instant examples

Recent arena outputs from GPT-5.5 Instant, picked from the highest-ranked matchups.

GPT-5.5 Instant license

GPT-5.5 Instant is a proprietary model available under its provider's product and API terms, has a knowledge cutoff of August 2025.

License
Proprietary
Hosted access
Knowledge cutoff
August 2025

Proprietary license - usage restrictions apply

GPT-5.5 Instant resources

Official sources for GPT-5.5 Instant: api documentation, official playground, official launch post.

GPT-5.5 Instant vs other models

The most-compared alternatives to GPT-5.5 Instant are Gemini 3 Pro, Seed 2.0 Lite, GPT-5.2. Open any pair side-by-side for benchmarks, pricing, context, and latency.

Models like GPT-5.5 Instant

Models ranked just above and below GPT-5.5 Instant by LLM Stats score.

 

Gemini 3 Pro

Score pending
 

Seed 2.0 Lite

Score pending
 

GPT-5.2

Score pending
 

Grok 4 Fast

Score pending
 

Qwen3 Max

Score pending
 

Qwen3.6 Plus

Score pending

FAQ

Common questions about GPT-5.5 Instant.

When was GPT-5.5 Instant released?

GPT-5.5 Instant was released on May 5, 2026 by OpenAI. This is the official GPT-5.5 Instant release date tracked on LLM Stats.

How much does GPT-5.5 Instant cost?

GPT-5.5 Instant costs $5.00 per million input tokens and $30.00 per million output tokens through the LLM Stats API, which works with any OpenAI-compatible SDK. Across tracked providers, the lowest price is $5.00 per million input tokens via OpenAI.

Is GPT-5.5 Instant available via API?

Yes. GPT-5.5 Instant is available through the LLM Stats API and works with any OpenAI-compatible SDK — point your client at the gateway base URL and pass the model name. It is served by 1 provider tracked on LLM Stats.

Who created GPT-5.5 Instant?

GPT-5.5 Instant was created by OpenAI.

What is the license for GPT-5.5 Instant?

GPT-5.5 Instant is released under the Proprietary license.

What is the knowledge cutoff date for GPT-5.5 Instant?

GPT-5.5 Instant has a knowledge cutoff of August 2025, meaning it was trained on data up to that point and may not know about events after it.

Is GPT-5.5 Instant multimodal?

Yes, GPT-5.5 Instant is multimodal and can accept both text and images as input.

What is GPT-5.5 Instant latency?

GPT-5.5 Instant p95 time to first token is 1.33 seconds via OpenAI over the trailing 7 days. Lower time to first token means the model begins responding sooner for chat, agents and API workloads.

Where can I use GPT-5.5 Instant?

GPT-5.5 Instant is available through 1 provider including OpenAI.

Where is the GPT-5.5 Instant paper or technical report?

GPT-5.5 Instant has a paper or technical report available at https://openai.com/index/gpt-5-5-instant/. Use that source for architecture, training, release and evaluation details.

What models should I compare GPT-5.5 Instant against?

Common GPT-5.5 Instant comparisons include GPT-5.5 Instant vs Gemini 3 Pro, GPT-5.5 Instant vs Seed 2.0 Lite, GPT-5.5 Instant vs GPT-5.2. Compare them side by side for benchmark scores, pricing, context window, latency and API availability.