- Organizations
- Gemma 4 12B
Gemma 4 12B: API Pricing, Context Window & Benchmarks
Gemma 4 12B is a language model from Google, released in May 2026, with multimodal input.
Gemma 4 12B is Google DeepMind's encoder-free multimodal instruction-tuned model with 11.95 billion parameters and a 256K context window. It supports text, image, audio, and video inputs with text output, projecting image patches and audio
Gemma 4 12B benchmarks
Rankings
Quality Tracker
Gemma 4 12B Performance Across Datasets
Scores sourced from the model's scorecard, paper, or official blog posts
Gemma 4 12B model size
Gemma 4 12B has 12.0 billion parameters. See how it compares to other models in the same parameter range.
Gemma 4 12B API
Available from the model provider
Gemma 4 12B has an official provider API. It is not currently routed through the LLM Stats gateway.
Read the official API documentationGemma 4 12B latency
Gemma 4 12B time to first token, sustained output throughput, and failed-request rate from live API traffic over the trailing 7 days.
Gemma 4 12B examples
Recent arena outputs from Gemma 4 12B, picked from the highest-ranked matchups.
Gemma 4 12B license
Gemma 4 12B is released under the Apache 2.0 license, which permits commercial use, has 12.0B parameters, has a knowledge cutoff of January 2025.
- License
- Apache 2.0
- Commercial use allowed
- Parameters
- 12.0B
- Knowledge cutoff
- January 2025
Apache License 2.0 - allows commercial use
Gemma 4 12B resources
Official sources for Gemma 4 12B: api documentation.
Gemma 4 12B vs other models
The most-compared alternatives to Gemma 4 12B are Claude 3.5 Sonnet, o1-pro, Qwen3 VL 235B A22B Instruct. Open any pair side-by-side for benchmarks, pricing, context, and latency.
Models like Gemma 4 12B
Models ranked just above and below Gemma 4 12B by LLM Stats score.
FAQ
Common questions about Gemma 4 12B.