← All comparisons

Molmo 7B-D vs Hermes 4 - Llama-3.1 405B (Non-reasoning)

Allen Institute for AI vs Nous Research — side-by-side benchmark comparison

Molmo 7B-DHermes 4 - Llama-3.1 405B (Non-reasoning)
Intelligence Index9.217.6
Coding Index1.218.1
Math Index0.015.3
Output speed (tok/s)0.040.8
Blended price ($/1M)$0.00$1.50
Time to first token (s)0.00s0.73s
aime
aime 250.0%15.3%
artificial analysis coding index1.2018.10
artificial analysis intelligence index9.2017.60
artificial analysis math index0.0%15.30
gpqa24.0%53.6%
hle5.1%4.2%
ifbench19.7%34.8%
lcr0.0%20.0%
livecodebench3.9%54.6%
math 500
mmlu pro37.1%72.9%
scicode3.6%34.6%
tau20.0%26.6%
terminalbench hard0.0%9.8%

Benchmark data from Artificial Analysis.