Mistral Large 3 675B Instruct 2512 — Benchmarks

Benchmark scores for Mistral Large 3 675B Instruct 2512 aggregated from public leaderboards, with how it ranks among open models. See hardware requirements for what you need to run it.

No overall rank for Mistral Large 3 675B Instruct 2512: an overall score is only published when a model has been measured on enough of the benchmarks current models are still submitted to. Its individual scores below stand on their own — see the methodology for how the overall score is built.

Coding

BenchmarkScoreOpen rankAll models
SciCode36.2%#33 / 62#127 / 161
LMArena WebDev1228.8#46 / 48#115 / 123

Instruction Following

BenchmarkScoreOpen rankAll models
LMArena Text1413.0#38 / 182#126 / 391

Reasoning

BenchmarkScoreOpen rankAll models
DTBench65.1%#36 / 67#147 / 210
LMCA16.7%#37 / 53#146 / 172
CritPt0.0%#55 / 62#161 / 172

Scores aggregated from public benchmark sources (each linked from the benchmark pages). llmrun does not run these benchmarks.