← All comparisons

Claude 4.5 Haiku (Non-reasoning) vs Devstral Small (May '25)

Anthropic vs Mistral — side-by-side benchmark comparison

Claude 4.5 Haiku (Non-reasoning)Devstral Small (May '25)
Intelligence Index31.018.0
Coding Index29.612.2
Math Index39.0
Output speed (tok/s)105.80.0
Blended price ($/1M)$2.19$0.00
Time to first token (s)0.65s0.00s
aime6.7%
aime 2539.0%
artificial analysis coding index29.6012.20
artificial analysis intelligence index31.0018.00
artificial analysis math index39.00
gpqa64.6%43.4%
hle4.3%4.0%
ifbench42.0%31.6%
lcr43.7%26.7%
livecodebench51.1%25.8%
math 50068.4%
mmlu pro80.0%63.2%
scicode34.4%24.5%
tau232.5%38.0%
terminalbench hard27.3%6.1%

Benchmark data from Artificial Analysis.