Qwen 1 8B — Benchmarks
Benchmark scores for Qwen 1 8B aggregated from public leaderboards, with how it ranks among open models. See hardware requirements for what you need to run it.
No overall rank for Qwen 1 8B: an overall score is only published when a model has been measured on enough of the benchmarks current models are still submitted to. Its individual scores below stand on their own — see the methodology for how the overall score is built.
Knowledge
| Benchmark | Score | Open rank | All models |
|---|---|---|---|
| ARC Challenge | 53.2% | #29 / 56 | #42 / 76 |
| BoolQ | 68.0% | #28 / 38 | #59 / 76 |
| MMLU | 28.2% | #80 / 85 | #130 / 136 |
Math
| Benchmark | Score | Open rank | All models |
|---|---|---|---|
| GSM8K | 21.2% | #54 / 62 | #73 / 93 |
Reasoning
| Benchmark | Score | Open rank | All models |
|---|---|---|---|
| BIG-Bench Hard | 28.2% | #38 / 38 | #50 / 50 |
| PIQA | 73.3% | #36 / 40 | #55 / 59 |
Scores aggregated from public benchmark sources (each linked from the benchmark pages). llmrun does not run these benchmarks.