- Organizations
- xAI
- Grok-1.5V
Grok-1.5V: Benchmarks, Pricing & Context Window
Grok-1.5V is a language model from xAI, released in April 2024, with multimodal input.
A multimodal model capable of processing text and visual information, including documents, diagrams, charts, screenshots, and photographs. Notable for strong real-world spatial understanding capabilities.
Grok-1.5V benchmarks
Capability tiers
Standing within each category, adjusted for leaderboard depth.
Real tasks performance
High-confidence performance for Grok-1.5V across real-world prompt categories. Only 95% intervals at most 4 points wide are shown.
Performance by conversation depth
How Grok-1.5V holds up as conversations get longer.
Quality Tracker
Grok-1.5V Performance Across Datasets
Scores sourced from the model's scorecard, paper, or official blog posts
Try now
Make it with
Grok-1.5V.
Grok-1.5V latency
Grok-1.5V time to first token, sustained output throughput, and failed-request rate from live model usage over the trailing 7 days.
Grok-1.5V examples
Recent arena outputs from Grok-1.5V, picked from the highest-ranked matchups.
Grok-1.5V license
Grok-1.5V is a proprietary model available under its provider's product and API terms.
- License
- Proprietary
- Hosted access
Proprietary license - usage restrictions apply
Grok-1.5V resources
Official sources for Grok-1.5V: provider documentation, official launch post.
Grok-1.5V vs other models
The most-compared alternatives to Grok-1.5V are Qwen3 VL 32B Thinking, DeepSeek VL2, Qwen3 VL 30B A3B Thinking. Open any pair side-by-side for benchmarks, pricing, context, and latency.
Models like Grok-1.5V
Models ranked just above and below Grok-1.5V by LLM Stats score.
FAQ
Common questions about Grok-1.5V.