Kimi K2.6 — Benchmarks
Benchmark scores for Kimi K2.6 aggregated from public leaderboards, with how it ranks among open models. See hardware requirements for what you need to run it.
Overall rank: #29 of 78 open modelscomposite 51.7/100 across 15 benchmarks in 4 categories · methodology
Coding
| Benchmark | Score | Open rank | All models |
|---|---|---|---|
| LiveBench Coding | 78.6 | #6 / 19 | #27 / 63 |
| SciCode | 53.5% | #3 / 62 | #48 / 161 |
| WeirdML | 55.9% | #7 / 46 | #57 / 161 |
| LMArena WebDev | 1508.6 | #12 / 48 | #43 / 123 |
Instruction Following
| Benchmark | Score | Open rank | All models |
|---|---|---|---|
| LMArena Text | 1460.4 | #8 / 182 | #52 / 391 |
Knowledge
| Benchmark | Score | Open rank | All models |
|---|---|---|---|
| SimpleQA | 34.9% | #8 / 15 | #45 / 80 |
Math
| Benchmark | Score | Open rank | All models |
|---|---|---|---|
| FrontierMath | 57.2% | #6 / 15 | #35 / 102 |
| FrontierMath Tier 4 | 25.6% | #5 / 11 | #38 / 64 |
| ProofBench | 16.0% | #13 / 24 | #45 / 68 |
| LiveBench Math | 84.3 | #14 / 19 | #52 / 63 |
| AIME 2024/2025 | 96.1% | #4 / 82 | #29 / 270 |
Reasoning
| Benchmark | Score | Open rank | All models |
|---|---|---|---|
| Chess Puzzles | 26.0% | #4 / 65 | #60 / 205 |
| Mystery Game Puzzles | 18.0% | #8 / 17 | #60 / 122 |
| DTBench | 90.9% | #5 / 67 | #60 / 210 |
| LMCA | 37.3% | #12 / 53 | #96 / 172 |
| GPQA Diamond | 90.8% | #8 / 94 | #33 / 289 |
| CritPt | 8.0% | #11 / 62 | #78 / 172 |
| LiveBench Reasoning | 79.4 | #12 / 19 | #49 / 63 |
Scores aggregated from public benchmark sources (each linked from the benchmark pages). llmrun does not run these benchmarks.