GPT OSS 20B — Benchmarks

Benchmark scores for GPT OSS 20B aggregated from public leaderboards, with how it ranks among open models. See hardware requirements for what you need to run it.

Overall rank: #67 of 123 open modelscomposite 44.9/100 across 6 benchmarks in 4 categories · methodology

Coding

BenchmarkScoreOpen rankAll models
Terminal-Bench3.4%#16 / 16#57 / 57
SciCode34.4%#36 / 50#132 / 155

Knowledge

BenchmarkScoreOpen rankAll models
MMLU-Pro73.6%#31 / 119#90 / 259

Math

BenchmarkScoreOpen rankAll models
AIME 2024/202565.3%#30 / 74#139 / 270

Reasoning

BenchmarkScoreOpen rankAll models
GPQA Diamond60.8%#38 / 83#189 / 291
CritPt1.4%#24 / 50#103 / 167

Scores aggregated from public benchmark sources (each linked from the benchmark pages). llmrun does not run these benchmarks.