Llama 4 Scout 17B 16E Instruct — Benchmarks

Benchmark scores for Llama 4 Scout 17B 16E Instruct aggregated from public leaderboards, with how it ranks among open models. See hardware requirements for what you need to run it.

Overall rank: #51 of 73 open modelscomposite 37.6/100 across 6 benchmarks in 3 categories · methodology

Knowledge

BenchmarkScoreOpen rankAll models
MMLU-Pro74.3%#29 / 119#88 / 259

Math

BenchmarkScoreOpen rankAll models
AIME 2024/20257.8%#22 / 34#120 / 155
MATH Level 562.3%#12 / 32#53 / 108
FrontierMath0.0%#12 / 12#101 / 101

Reasoning

BenchmarkScoreOpen rankAll models
GPQA Diamond51.8%#19 / 46#116 / 182
ARC-AGI0.5%#10 / 10#157 / 158

Scores aggregated from public benchmark sources (each linked from the benchmark pages). llmrun does not run these benchmarks.