Llama 2 7B HF — Benchmarks

Benchmark scores for Llama 2 7B HF aggregated from public leaderboards, with how it ranks among open models. See hardware requirements for what you need to run it.

Overall rank: #59 of 73 open modelscomposite 30.7/100 across 5 benchmarks in 3 categories · methodology

Knowledge

BenchmarkScoreOpen rankAll models
HellaSwag77.2%#18 / 42#37 / 76
MMLU-Pro20.3%#108 / 119#244 / 259
MMLU45.8%#61 / 76#113 / 136

Math

BenchmarkScoreOpen rankAll models
GSM8K16.7%#53 / 59#76 / 93

Reasoning

BenchmarkScoreOpen rankAll models
BIG-Bench Hard39.2%#28 / 37#38 / 50

Scores aggregated from public benchmark sources (each linked from the benchmark pages). llmrun does not run these benchmarks.