Comparison

GPT-4 Multilingual Performance on MMLU

On translated versions of the MMLU benchmark, GPT-4 demonstrates robust cross-lingual understanding by outperforming the English-language benchmark scores of preceding state-of-the-art models (such as Chinchilla and PaLM) across the vast majority of evaluated languages. Notably, GPT-4 surpasses prior English-language models even in low-resource languages with limited pre-training data availability, such as Latvian, Welsh, and Swahili.

0

1

Updated 2026-09-07

Tags

Prep Sessions

Transformer Architecture and Large Language Model Capabilities @ University of Michigan - Ann Arbor

Ch.2 Model Scaling and Capability Evaluation - Transformer Architecture and Large Language Model Capabilities @ University of Michigan - Ann Arbor

Multilingual Language Understanding on MMLU - Transformer Architecture and Large Language Model Capabilities @ University of Michigan - Ann Arbor