← All comparisons

Claude 4.5 Haiku (Reasoning) vs Phi-4 Multimodal Instruct

Anthropic vs Microsoft — side-by-side benchmark comparison

Claude 4.5 Haiku (Reasoning)Phi-4 Multimodal Instruct
Intelligence Index37.110.0
Coding Index32.6
Math Index83.7
Output speed (tok/s)142.216.6
Blended price ($/1M)$2.19$0.00
Time to first token (s)10.48s1.33s
aime9.3%
aime 2583.7%
artificial analysis coding index32.60
artificial analysis intelligence index37.1010.00
artificial analysis math index83.70
gpqa67.2%31.5%
hle9.7%4.4%
ifbench54.3%
lcr70.3%
livecodebench61.5%13.1%
math 50069.3%
mmlu pro76.0%48.5%
scicode43.3%11.0%
tau254.7%
terminalbench hard27.3%

Benchmark data from Artificial Analysis.