← All comparisons

Claude 4.5 Haiku (Reasoning) vs Devstral Small (Jul '25)

Anthropic vs Mistral — side-by-side benchmark comparison

Claude 4.5 Haiku (Reasoning)Devstral Small (Jul '25)
Intelligence Index37.115.2
Coding Index32.612.1
Math Index83.729.3
Output speed (tok/s)142.2183.4
Blended price ($/1M)$2.19$0.15
Time to first token (s)10.48s0.40s
aime0.3%
aime 2583.7%29.3%
artificial analysis coding index32.6012.10
artificial analysis intelligence index37.1015.20
artificial analysis math index83.7029.30
gpqa67.2%41.4%
hle9.7%3.7%
ifbench54.3%34.6%
lcr70.3%17.0%
livecodebench61.5%25.4%
math 50063.5%
mmlu pro76.0%62.2%
scicode43.3%24.3%
tau254.7%28.4%
terminalbench hard27.3%6.1%

Benchmark data from Artificial Analysis.