The AI arena is free today

Open Superagent

Beam 128K

Progress Over Time

Interactive timeline showing model performance evolution on Beam 128K

State-of-the-art frontier
Open
Proprietary

Beam 128K Leaderboard

1 models
ContextCostLicense
130B
Notice missing or incorrect data?
About this benchmark

What is Beam 128K?

Beam 128K evaluates reasoning over long inputs at a 128K-token context length.

Beam 128K is a text benchmark evaluating models on long context and reasoning tasks. LLM Stats tracks 1 models on this benchmark, scored on a 0–1 scale. The current average is 0.7, with the leader at 0.7.

Compare leaders on the best AI for long context and best AI for reasoning leaderboards.

Current leaders

Muse Glimmer-30B from Meta currently leads the Beam 128K leaderboard with a score of 0.651 across 1 evaluated AI models.

FAQ

Common questions about the Beam 128K benchmark and leaderboard.

What is the Beam 128K benchmark?

Beam 128K evaluates reasoning over long inputs at a 128K-token context length.

What is the Beam 128K leaderboard?

The Beam 128K leaderboard ranks 1 AI models based on their performance on this benchmark. Currently, Muse Glimmer-30B by Meta leads with a score of 0.651. The average score across all models is 0.651.

What is the highest Beam 128K score?

The highest Beam 128K score is 0.651, achieved by Muse Glimmer-30B from Meta.

How many models are evaluated on Beam 128K?

1 models have been evaluated on the Beam 128K benchmark, with 0 verified results and 1 self-reported results.

What categories does Beam 128K cover?

Beam 128K is categorized under long context and reasoning. The benchmark evaluates text models.

What is the best open-source model on Beam 128K?

Muse Glimmer-30B by Meta is the top-ranked open-source model on Beam 128K, with a score of 0.651 (rank #1).

How recent are the Beam 128K leaderboard results?

The Beam 128K leaderboard was last updated in August 2026 and currently includes 1 evaluated models.