The AI arena is free today

Open Superagent

MEGA XCOPA

Paper

Progress Over Time

Interactive timeline showing model performance evolution on MEGA XCOPA

State-of-the-art frontier
Open
Proprietary

MEGA XCOPA Leaderboard

2 models
ContextCostLicense
160B
24B
Notice missing or incorrect data?
About this benchmark

What is MEGA XCOPA?

XCOPA (Cross-lingual Choice of Plausible Alternatives) as part of the MEGA benchmark suite. A typologically diverse multilingual dataset for causal commonsense reasoning in 11 languages, including resource-poor languages like Eastern Apurímac Quechua and Haitian Creole. Requires models to select which choice is the effect or cause of a given premise.

MEGA XCOPA is a text benchmark evaluating models on language and reasoning tasks. LLM Stats tracks 2 models on this benchmark, scored on a 0–1 scale. The current average is 0.7, with the leader at 0.8.

Compare leaders on the best AI for language and best AI for reasoning leaderboards.

Current leaders

Phi-3.5-MoE-instruct from Microsoft currently leads the MEGA XCOPA leaderboard with a score of 0.766 across 2 evaluated AI models.

1Phi-3.5-MoE-instructMicrosoft76.6%
2Phi-3.5-mini-instructMicrosoft63.1%

FAQ

Common questions about the MEGA XCOPA benchmark and leaderboard.

What is the MEGA XCOPA benchmark?

XCOPA (Cross-lingual Choice of Plausible Alternatives) as part of the MEGA benchmark suite. A typologically diverse multilingual dataset for causal commonsense reasoning in 11 languages, including resource-poor languages like Eastern Apurímac Quechua and Haitian Creole. Requires models to select which choice is the effect or cause of a given premise.

What is the MEGA XCOPA leaderboard?

The MEGA XCOPA leaderboard ranks 2 AI models based on their performance on this benchmark. Currently, Phi-3.5-MoE-instruct by Microsoft leads with a score of 0.766. The average score across all models is 0.699.

What is the highest MEGA XCOPA score?

The highest MEGA XCOPA score is 0.766, achieved by Phi-3.5-MoE-instruct from Microsoft.

How many models are evaluated on MEGA XCOPA?

2 models have been evaluated on the MEGA XCOPA benchmark, with 0 verified results and 2 self-reported results.

Where can I find the MEGA XCOPA paper?

The MEGA XCOPA paper is available at https://arxiv.org/abs/2005.00333. The paper details the methodology, dataset construction, and evaluation criteria.

What categories does MEGA XCOPA cover?

MEGA XCOPA is categorized under language and reasoning. The benchmark evaluates text models with multilingual support.

What is the best open-source model on MEGA XCOPA?

Phi-3.5-MoE-instruct by Microsoft is the top-ranked open-source model on MEGA XCOPA, with a score of 0.766 (rank #1).

How recent are the MEGA XCOPA leaderboard results?

The MEGA XCOPA leaderboard was last updated in August 2026 and currently includes 2 evaluated models.