Llama 2 70B HF — Benchmarks

Benchmark scores for Llama 2 70B HF aggregated from public leaderboards, with how it ranks among open models. See hardware requirements for what you need to run it.

Overall rank: #11 of 73 open modelscomposite 63.8/100 across 5 benchmarks in 3 categories · methodology

Knowledge

BenchmarkScoreOpen rankAll models
HellaSwag85.3%#6 / 42#10 / 76
MMLU-Pro37.5%#87 / 119#208 / 259
MMLU69.9%#26 / 76#62 / 136

Math

BenchmarkScoreOpen rankAll models
GSM8K69.6%#17 / 59#26 / 93

Reasoning

BenchmarkScoreOpen rankAll models
BIG-Bench Hard64.9%#9 / 37#13 / 50

Scores aggregated from public benchmark sources (each linked from the benchmark pages). llmrun does not run these benchmarks.