← All comparisons

Claude 4.5 Haiku (Reasoning) vs DeepHermes 3 - Llama-3.1 8B Preview (Non-reasoning)

Anthropic vs Nous Research — side-by-side benchmark comparison

Claude 4.5 Haiku (Reasoning)DeepHermes 3 - Llama-3.1 8B Preview (Non-reasoning)
Intelligence Index37.17.6
Coding Index32.6
Math Index83.7
Output speed (tok/s)142.20.0
Blended price ($/1M)$2.19$0.00
Time to first token (s)10.48s0.00s
aime0.0%
aime 2583.7%
artificial analysis coding index32.60
artificial analysis intelligence index37.107.60
artificial analysis math index83.70
gpqa67.2%27.0%
hle9.7%4.3%
ifbench54.3%
lcr70.3%
livecodebench61.5%8.5%
math 50021.8%
mmlu pro76.0%36.5%
scicode43.3%9.1%
tau254.7%
terminalbench hard27.3%

Benchmark data from Artificial Analysis.