← All comparisons

Claude 4.5 Haiku (Reasoning) vs Devstral Small (May '25)

Anthropic vs Mistral — side-by-side benchmark comparison

Claude 4.5 Haiku (Reasoning)Devstral Small (May '25)
Intelligence Index37.118.0
Coding Index32.612.2
Math Index83.7
Output speed (tok/s)142.20.0
Blended price ($/1M)$2.19$0.00
Time to first token (s)10.48s0.00s
aime6.7%
aime 2583.7%
artificial analysis coding index32.6012.20
artificial analysis intelligence index37.1018.00
artificial analysis math index83.70
gpqa67.2%43.4%
hle9.7%4.0%
ifbench54.3%31.6%
lcr70.3%26.7%
livecodebench61.5%25.8%
math 50068.4%
mmlu pro76.0%63.2%
scicode43.3%24.5%
tau254.7%38.0%
terminalbench hard27.3%6.1%

Benchmark data from Artificial Analysis.