← All comparisons

Claude 4.5 Haiku (Reasoning) vs Phi-4

Anthropic vs Microsoft — side-by-side benchmark comparison

Claude 4.5 Haiku (Reasoning)Phi-4
Intelligence Index37.110.4
Coding Index32.611.2
Math Index83.718.0
Output speed (tok/s)142.241.1
Blended price ($/1M)$2.19$0.22
Time to first token (s)10.48s0.50s
aime14.3%
aime 2583.7%18.0%
artificial analysis coding index32.6011.20
artificial analysis intelligence index37.1010.40
artificial analysis math index83.7018.00
gpqa67.2%57.5%
hle9.7%4.1%
ifbench54.3%23.5%
lcr70.3%0.0%
livecodebench61.5%23.1%
math 50081.0%
mmlu pro76.0%71.4%
scicode43.3%26.0%
tau254.7%0.0%
terminalbench hard27.3%3.8%

Benchmark data from Artificial Analysis.