Qwen3.5 9B — Benchmarks

Benchmark scores for Qwen3.5 9B aggregated from public leaderboards, with how it ranks among open models. See hardware requirements for what you need to run it.

Overall rank: #35 of 78 open modelscomposite 48/100 across 9 benchmarks in 4 categories · methodology

Coding

BenchmarkScoreOpen rankAll models
SciCode27.6%#51 / 62#147 / 161
Terminal-Bench9.2%#15 / 16#56 / 57

Knowledge

BenchmarkScoreOpen rankAll models
MMLU-Pro82.5%#17 / 144#53 / 257

Math

BenchmarkScoreOpen rankAll models
AIME 2024/202561.7%#35 / 82#151 / 270

Reasoning

BenchmarkScoreOpen rankAll models
Chess Puzzles12.0%#23 / 65#128 / 205
DTBench71.2%#31 / 67#133 / 210
LMCA24.5%#28 / 53#128 / 172
GPQA Diamond79.0%#24 / 94#116 / 289
CritPt0.3%#41 / 62#133 / 172

Scores aggregated from public benchmark sources (each linked from the benchmark pages). llmrun does not run these benchmarks.