The AI arena is free today

Open Superagent
GoogleReleased on Feb 19, 2026

Gemini 3.1 Pro: Benchmarks, Pricing & Context Window

Gemini 3.1 Pro is a language model from Google, released in February 2026, with multimodal input, a 1.0M-token context window, and pricing from $2.00/M input and $12.00/M output.

Gemini 3.1 Pro is the latest model in the Gemini 3 series. It excels at complex tasks requiring broad world knowledge and advanced reasoning across modalities. Gemini 3.1 Pro uses dynamic thinking by default to reason through prompts, and

Input
TextImageAudioVideo
Output
Text

Gemini 3.1 Pro benchmarks

Capability tiers

Standing within each category, adjusted for leaderboard depth.

Real tasks performance

High-confidence performance for Gemini 3.1 Pro across real-world prompt categories. Only 95% intervals at most 4 points wide are shown.

Performance by conversation depth

How Gemini 3.1 Pro holds up as conversations get longer.

Quality Tracker

Gemini 3.1 Pro Performance Across Datasets

Scores sourced from the model's scorecard, paper, or official blog posts

LLM Stats Logollm-stats.com - Mon Oct 05 2026
Notice missing or incorrect data?

Gemini 3.1 Pro pricing

Providers

Gemini 3.1 Pro starts at $2.00 per million input tokens and $12.00 per million output tokens via DeepInfra. See all 2 providers below with their per-token pricing, latency, throughput, and modality support.

ProviderInput $/MCached input $/MOutput $/MContext in / outTTFT p95 sOutput p5 c/sModalities in / out
DeepInfra logoDeepInfra
$2.00—$12.001.0M/1.0M
—
—
/
Google logoGoogle
$2.50—$15.001.0M/65.5K
0.60
—
/

Cached input is the discounted price for prompt tokens served from a provider cache. TTFT is time to first token. Output is characters per second; p5 is the sustained floor exceeded by 95% of observed requests.

Loading chart...
Loading chart...
Loading chart...

Gemini 3.1 Pro context window

Input and output token limits for Gemini 3.1 Pro, plus how it ranks on long-context understanding.

InputOutput
1.0Mtokens
1Mtokens
≈ 1.6k pages of text
1.0M
8K128K1M

Try now

huggle
Gemini 3.1 Proin Huggle

Make it with
Gemini 3.1 Pro.

Gemini 3.1 Pro

Gemini 3.1 Pro latency

Gemini 3.1 Pro time to first token, sustained output throughput, and failed-request rate from live model usage over the trailing 7 days.

Gemini 3.1 Pro examples

Recent arena outputs from Gemini 3.1 Pro, picked from the highest-ranked matchups.

Gemini 3.1 Pro license

Gemini 3.1 Pro is a proprietary model available under its provider's product and API terms, has a knowledge cutoff of January 2025.

License
Proprietary
Hosted access
Knowledge cutoff
January 2025

Proprietary license - usage restrictions apply

Gemini 3.1 Pro resources

Official sources for Gemini 3.1 Pro: provider documentation, official playground, official launch post.

Gemini 3.1 Pro vs other models

The most-compared alternatives to Gemini 3.1 Pro are GLM-5.1, Claude Opus 4.6, Gemini 3 Pro. Open any pair side-by-side for benchmarks, pricing, context, and latency.

Models like Gemini 3.1 Pro

Models ranked just above and below Gemini 3.1 Pro by LLM Stats score.

 

GLM-5.1

Score pending
 

Claude Opus 4.6

Score pending
 

Gemini 3 Pro

Score pending
 

Claude Mythos Preview

Score pending
 

Grok-4 Heavy

Score pending
 

Qwen3.7 Max

Score pending

FAQ

Common questions about Gemini 3.1 Pro.

When was Gemini 3.1 Pro released?

Gemini 3.1 Pro was released on February 19, 2026 by Google. This is the official Gemini 3.1 Pro release date tracked on LLM Stats.

How much does Gemini 3.1 Pro cost?

Gemini 3.1 Pro pricing starts at $2.00 per million input tokens and $12.00 per million output tokens via DeepInfra, the lowest price among tracked providers.

Who created Gemini 3.1 Pro?

Gemini 3.1 Pro was created by Google.

What is the license for Gemini 3.1 Pro?

Gemini 3.1 Pro is released under the Proprietary license.

What is the knowledge cutoff date for Gemini 3.1 Pro?

Gemini 3.1 Pro has a knowledge cutoff of January 2025, meaning it was trained on data up to that point and may not know about events after it.

Is Gemini 3.1 Pro multimodal?

Yes, Gemini 3.1 Pro is multimodal and can accept both text and images as input.

What is Gemini 3.1 Pro latency?

Gemini 3.1 Pro p95 time to first token is 0.60 seconds via Google over the trailing 7 days. Lower time to first token means the model begins responding sooner for chat, agents and model workloads.

Where can I use Gemini 3.1 Pro?

Gemini 3.1 Pro is available through 2 providers including DeepInfra, Google.

Where is the Gemini 3.1 Pro paper or technical report?

Gemini 3.1 Pro has a paper or technical report available at https://deepmind.google/models/evals-methodology/gemini-3-1-pro. Use that source for architecture, training, release and evaluation details.

What models should I compare Gemini 3.1 Pro against?

Common Gemini 3.1 Pro comparisons include Gemini 3.1 Pro vs GLM-5.1, Gemini 3.1 Pro vs Claude Opus 4.6, Gemini 3.1 Pro vs Gemini 3 Pro. Compare them side by side for benchmark scores, pricing, context window, latency and provider availability.