Mistral Small 3.2 24B Instruct 2506 — Benchmarks

Benchmark scores for Mistral Small 3.2 24B Instruct 2506 aggregated from public leaderboards, with how it ranks among open models. See hardware requirements for what you need to run it.

Overall rank: #69 of 78 open modelscomposite 24.3/100 across 6 benchmarks in 3 categories · methodology

Coding

BenchmarkScoreOpen rankAll models
SciCode26.4%#53 / 62#150 / 161

Instruction Following

BenchmarkScoreOpen rankAll models
LMArena Text1356.7#66 / 182#187 / 391

Math

BenchmarkScoreOpen rankAll models
AIME 2024/202530.3%#47 / 82#202 / 270

Reasoning

BenchmarkScoreOpen rankAll models
Chess Puzzles1.0%#44 / 65#180 / 205
DTBench59.9%#48 / 67#168 / 210
GPQA Diamond49.0%#53 / 94#216 / 289
CritPt0.0%#57 / 62#163 / 172

Scores aggregated from public benchmark sources (each linked from the benchmark pages). llmrun does not run these benchmarks.