Hermes 2 Theta Llama 3 70B — Benchmarks
Benchmark scores for Hermes 2 Theta Llama 3 70B aggregated from public leaderboards, with how it ranks among open models. See hardware requirements for what you need to run it.
No overall rank for Hermes 2 Theta Llama 3 70B: an overall score is only published when a model has been measured on enough of the benchmarks current models are still submitted to. Its individual scores below stand on their own — see the methodology for how the overall score is built.
Math
| Benchmark | Score | Open rank | All models |
|---|---|---|---|
| MATH Level 5 | 22.7% | #29 / 39 | #87 / 107 |
| AIME 2024/2025 | 2.5% | #72 / 82 | #251 / 270 |
Reasoning
| Benchmark | Score | Open rank | All models |
|---|---|---|---|
| GPQA Diamond | 37.5% | #71 / 94 | #251 / 289 |
Scores aggregated from public benchmark sources (each linked from the benchmark pages). llmrun does not run these benchmarks.