← All comparisons

Claude 4.5 Haiku (Non-reasoning) vs Devstral Small (Jul '25)

Anthropic vs Mistral — side-by-side benchmark comparison

Claude 4.5 Haiku (Non-reasoning)Devstral Small (Jul '25)
Intelligence Index31.015.2
Coding Index29.612.1
Math Index39.029.3
Output speed (tok/s)105.8183.4
Blended price ($/1M)$2.19$0.15
Time to first token (s)0.65s0.40s
aime0.3%
aime 2539.0%29.3%
artificial analysis coding index29.6012.10
artificial analysis intelligence index31.0015.20
artificial analysis math index39.0029.30
gpqa64.6%41.4%
hle4.3%3.7%
ifbench42.0%34.6%
lcr43.7%17.0%
livecodebench51.1%25.4%
math 50063.5%
mmlu pro80.0%62.2%
scicode34.4%24.3%
tau232.5%28.4%
terminalbench hard27.3%6.1%

Benchmark data from Artificial Analysis.