← All comparisons

DeepHermes 3 - Llama-3.1 8B Preview (Non-reasoning) vs OLMo 2 32B

Nous Research vs Allen Institute for AI — side-by-side benchmark comparison

DeepHermes 3 - Llama-3.1 8B Preview (Non-reasoning)OLMo 2 32B
Intelligence Index7.610.6
Coding Index2.7
Math Index3.3
Output speed (tok/s)0.00.0
Blended price ($/1M)$0.00$0.00
Time to first token (s)0.00s0.00s
aime0.0%
aime 253.3%
artificial analysis coding index2.70
artificial analysis intelligence index7.6010.60
artificial analysis math index3.30
gpqa27.0%32.8%
hle4.3%3.7%
ifbench38.1%
lcr0.0%
livecodebench8.5%6.8%
math 50021.8%
mmlu pro36.5%51.1%
scicode9.1%8.0%
tau20.0%
terminalbench hard0.0%

Benchmark data from Artificial Analysis.