The AI arena is free today

Open Superagent
OpenAIReleased on Aug 5, 2025

GPT OSS 20B High: API Pricing, Context Window & Benchmarks

GPT OSS 20B High is a language model from OpenAI, released in August 2025.

GPT-OSS-20B High provides enhanced reasoning capabilities with high-effort thinking for complex problems. This variant offers deeper analysis and more thorough responses compared to the base model, making it ideal for challenging tasks

Input
Text
Output
Text

GPT OSS 20B High benchmarks

Capability tiers

Standing within each category, adjusted for leaderboard depth.

Real tasks performance

How GPT OSS 20B High performs across real-world prompt categories.

Performance by conversation depth

How GPT OSS 20B High holds up as conversations get longer.

Quality Tracker

GPT OSS 20B High Performance Across Datasets

Scores sourced from the model's scorecard, paper, or official blog posts

LLM Stats Logollm-stats.com - Mon Aug 24 2026
Notice missing or incorrect data?

GPT OSS 20B High pricing

Providers

GPT OSS 20B High starts at $0.100 per million input tokens and $0.500 per million output tokens via OpenAI.

ProviderInput $/MCached input $/MOutput $/MContext in / outTTFT p95 sOutput p5 c/sModalities in / out
OpenAI logoOpenAI
$0.100$0.500131.1K/131.1K
6.50
/

Cached input is the discounted price for prompt tokens served from a provider cache. TTFT is time to first token. Output is characters per second; p5 is the sustained floor exceeded by 95% of observed requests.

GPT OSS 20B High model size

GPT OSS 20B High has 20.9 billion parameters. See how it compares to other models in the same parameter range.

Parameters
20.9B
Medium (10–30B)
20.9B
1B7B70B405B

GPT OSS 20B High context window

Input and output token limits for GPT OSS 20B High, plus how it ranks on long-context understanding.

InputOutput
131Ktokens
131Ktokens
197 pages of text
131K
8K128K1M

GPT OSS 20B High latency

GPT OSS 20B High time to first token, sustained output throughput, and failed-request rate from live API traffic over the trailing 7 days.

GPT OSS 20B High examples

Recent arena outputs from GPT OSS 20B High, picked from the highest-ranked matchups.

GPT OSS 20B High license

GPT OSS 20B High is released under the Apache 2.0 license, which permits commercial use, has 20.9B parameters.

License
Apache 2.0
Commercial use allowed
Parameters
20.9B

Apache License 2.0 - allows commercial use

GPT OSS 20B High resources

Official sources for GPT OSS 20B High: official playground, paper or system card, official launch post, source repository, model weights.

GPT OSS 20B High vs other models

The most-compared alternatives to GPT OSS 20B High are GPT-5.1 Medium, Seed 2.0 Pro, LongCat-Flash-Thinking-2601. Open any pair side-by-side for benchmarks, pricing, context, and latency.

Models like GPT OSS 20B High

Models ranked just above and below GPT OSS 20B High by LLM Stats score.

 

GPT-5.1 Medium

Score pending
 

Seed 2.0 Pro

Score pending
 

LongCat-Flash-Thinking-2601

Score pending
 

Gemini 2.0 Flash Thinking

Score pending
 

Qwen3 VL 30B A3B Thinking

Score pending
 

Nemotron 3 Nano (30B A3B)

Score pending

FAQ

Common questions about GPT OSS 20B High.

When was GPT OSS 20B High released?

GPT OSS 20B High was released on August 5, 2025 by OpenAI. This is the official GPT OSS 20B High release date tracked on LLM Stats.

How much does GPT OSS 20B High cost?

GPT OSS 20B High pricing starts at $0.10 per million input tokens and $0.50 per million output tokens via OpenAI, the lowest price among tracked providers.

How big is GPT OSS 20B High?

GPT OSS 20B High has 20.9 billion parameters. It ships as an open-weight model, so you can download and run it on your own hardware.

Who created GPT OSS 20B High?

GPT OSS 20B High was created by OpenAI.

What is the license for GPT OSS 20B High?

GPT OSS 20B High is released under the Apache 2.0 license. This is an open-source / open-weight license that permits self-hosting.

What is GPT OSS 20B High latency?

GPT OSS 20B High p95 time to first token is 6.50 seconds via OpenAI over the trailing 7 days. Lower time to first token means the model begins responding sooner for chat, agents and API workloads.

Where can I use GPT OSS 20B High?

GPT OSS 20B High is available through 1 provider including OpenAI.

Where is the GPT OSS 20B High paper or technical report?

GPT OSS 20B High has a paper or technical report available at https://cdn.openai.com/pdf/419b6906-9da6-406c-a19d-1bb078ac7637/oai_gpt-oss_model_card.pdf. Use that source for architecture, training, release and evaluation details.

What models should I compare GPT OSS 20B High against?

Common GPT OSS 20B High comparisons include GPT OSS 20B High vs GPT-5.1 Medium, GPT OSS 20B High vs Seed 2.0 Pro, GPT OSS 20B High vs LongCat-Flash-Thinking-2601. Compare them side by side for benchmark scores, pricing, context window, latency and API availability.