GDP.pdf
Progress Over Time
Interactive timeline showing model performance evolution on GDP.pdf
GDP.pdf Leaderboard
| Context | Cost | License | ||||
|---|---|---|---|---|---|---|
| 1 | Anthropic | — | 1.0M | $3.00 / $15.00 | ||
| 2 | OpenAI | — | 1.1M | $5.00 / $30.00 | ||
| 3 | Anthropic | — | 1.0M | $10.00 / $50.00 | ||
| 4 | OpenAI | — | 1.1M | $2.50 / $15.00 | ||
| 5 | OpenAI | — | 1.1M | $1.00 / $6.00 |
What is GDP.pdf?
GDP.pdf is a knowledge-work vision benchmark that evaluates models on economically valuable professional tasks presented as visual documents (PDFs), testing document-based reasoning, chart and table interpretation, and problem solving without tools.
GDP.pdf is a multimodal benchmark evaluating models on multimodal, reasoning, general, and vision tasks. LLM Stats tracks 5 models on this benchmark, scored on a 0–1 scale. The current average is 0.4, with the leader at 0.8.
Compare leaders on the best AI for multimodal, best AI for reasoning, best AI for general and best AI for vision leaderboards.
Current leaders
Claude Sonnet 5 from Anthropic currently leads the GDP.pdf leaderboard with a score of 0.816 across 5 evaluated AI models.
FAQ
Common questions about the GDP.pdf benchmark and leaderboard.