The AI arena is free today

Open Superagent

MultiLF

Progress Over Time

Interactive timeline showing model performance evolution on MultiLF

State-of-the-art frontier
Open
Proprietary

MultiLF Leaderboard

2 models
ContextCostLicense
1
Alibaba Cloud / Qwen Team
Alibaba Cloud / Qwen Team
33B128K$0.10 / $0.30
2
Alibaba Cloud / Qwen Team
Alibaba Cloud / Qwen Team
235B
Notice missing or incorrect data?
About this benchmark

What is MultiLF?

MultiLF benchmark

MultiLF is a text benchmark evaluating models on general tasks. LLM Stats tracks 2 models on this benchmark, scored on a 0–1 scale. The current average is 0.7, with the leader at 0.7.

Compare leaders on the best AI for general leaderboards.

Current leaders

Qwen3 32B from Alibaba Cloud / Qwen Team currently leads the MultiLF leaderboard with a score of 0.730 across 2 evaluated AI models.

1Qwen3 32BAlibaba Cloud / Qwen Team73.0%
2Qwen3 235B A22BAlibaba Cloud / Qwen Team71.9%

FAQ

Common questions about the MultiLF benchmark and leaderboard.

What is the MultiLF benchmark?

MultiLF benchmark

What is the MultiLF leaderboard?

The MultiLF leaderboard ranks 2 AI models based on their performance on this benchmark. Currently, Qwen3 32B by Alibaba Cloud / Qwen Team leads with a score of 0.730. The average score across all models is 0.724.

What is the highest MultiLF score?

The highest MultiLF score is 0.730, achieved by Qwen3 32B from Alibaba Cloud / Qwen Team.

How many models are evaluated on MultiLF?

2 models have been evaluated on the MultiLF benchmark, with 0 verified results and 2 self-reported results.

What categories does MultiLF cover?

MultiLF is categorized under general. The benchmark evaluates text models.

What is the best open-source model on MultiLF?

Qwen3 32B by Alibaba Cloud / Qwen Team is the top-ranked open-source model on MultiLF, with a score of 0.730 (rank #1).

Which model offers the best value on MultiLF?

Among models scoring within 10% of the leader, Qwen3 32B from Alibaba Cloud / Qwen Team is the cheapest, at $0.10 per million input tokens with a score of 0.730.

How recent are the MultiLF leaderboard results?

The MultiLF leaderboard was last updated in August 2026 and currently includes 2 evaluated models.
MultiLF Leaderboard