← All comparisons

DeepHermes 3 - Llama-3.1 8B Preview (Non-reasoning) vs OLMo 2 7B

Nous Research vs Allen Institute for AI — side-by-side benchmark comparison

DeepHermes 3 - Llama-3.1 8B Preview (Non-reasoning)OLMo 2 7B
Intelligence Index7.69.3
Coding Index1.2
Math Index0.7
Output speed (tok/s)0.00.0
Blended price ($/1M)$0.00$0.00
Time to first token (s)0.00s0.00s
aime0.0%
aime 250.7%
artificial analysis coding index1.20
artificial analysis intelligence index7.609.30
artificial analysis math index70.0%
gpqa27.0%28.8%
hle4.3%5.5%
ifbench24.4%
lcr0.0%
livecodebench8.5%4.1%
math 50021.8%
mmlu pro36.5%28.2%
scicode9.1%3.7%
tau20.0%
terminalbench hard0.0%

Benchmark data from Artificial Analysis.