Hermes 2 Theta Llama 3 70B — Benchmarks

Benchmark scores for Hermes 2 Theta Llama 3 70B aggregated from public leaderboards, with how it ranks among open models. See hardware requirements for what you need to run it.

No overall rank for Hermes 2 Theta Llama 3 70B: an overall score is only published when a model has been measured on enough of the benchmarks current models are still submitted to. Its individual scores below stand on their own — see the methodology for how the overall score is built.

Math

BenchmarkScoreOpen rankAll models
MATH Level 522.7%#29 / 39#87 / 107
AIME 2024/20252.5%#72 / 82#251 / 270

Reasoning

BenchmarkScoreOpen rankAll models
GPQA Diamond37.5%#71 / 94#251 / 289

Scores aggregated from public benchmark sources (each linked from the benchmark pages). llmrun does not run these benchmarks.