Beam 128K
Progress Over Time
Interactive timeline showing model performance evolution on Beam 128K
State-of-the-art frontier
Open
Proprietary
Beam 128K Leaderboard
1 models
| Context | Cost | License | ||||
|---|---|---|---|---|---|---|
| 1 | Meta | 30B | — | — |
Notice missing or incorrect data?
What is Beam 128K?
Beam 128K evaluates reasoning over long inputs at a 128K-token context length.
Beam 128K is a text benchmark evaluating models on long context and reasoning tasks. LLM Stats tracks 1 models on this benchmark, scored on a 0–1 scale. The current average is 0.7, with the leader at 0.7.
Compare leaders on the best AI for long context and best AI for reasoning leaderboards.
Current leaders
Muse Glimmer-30B from Meta currently leads the Beam 128K leaderboard with a score of 0.651 across 1 evaluated AI models.
FAQ
Common questions about the Beam 128K benchmark and leaderboard.