← All comparisons

Claude 4.5 Haiku (Reasoning) vs Phi-4 Mini Instruct

Anthropic vs Microsoft — side-by-side benchmark comparison

Claude 4.5 Haiku (Reasoning)Phi-4 Mini Instruct
Intelligence Index37.18.4
Coding Index32.63.6
Math Index83.76.7
Output speed (tok/s)142.20.0
Blended price ($/1M)$2.19$0.00
Time to first token (s)10.48s0.00s
aime3.0%
aime 2583.7%6.7%
artificial analysis coding index32.603.60
artificial analysis intelligence index37.108.40
artificial analysis math index83.706.70
gpqa67.2%33.1%
hle9.7%4.2%
ifbench54.3%21.1%
lcr70.3%13.7%
livecodebench61.5%12.6%
math 50069.6%
mmlu pro76.0%46.5%
scicode43.3%10.8%
tau254.7%8.2%
terminalbench hard27.3%0.0%

Benchmark data from Artificial Analysis.