← All comparisons

Claude 4.5 Haiku (Non-reasoning) vs Phi-4

Anthropic vs Microsoft — side-by-side benchmark comparison

Claude 4.5 Haiku (Non-reasoning)Phi-4
Intelligence Index31.010.4
Coding Index29.611.2
Math Index39.018.0
Output speed (tok/s)105.841.1
Blended price ($/1M)$2.19$0.22
Time to first token (s)0.65s0.50s
aime14.3%
aime 2539.0%18.0%
artificial analysis coding index29.6011.20
artificial analysis intelligence index31.0010.40
artificial analysis math index39.0018.00
gpqa64.6%57.5%
hle4.3%4.1%
ifbench42.0%23.5%
lcr43.7%0.0%
livecodebench51.1%23.1%
math 50081.0%
mmlu pro80.0%71.4%
scicode34.4%26.0%
tau232.5%0.0%
terminalbench hard27.3%3.8%

Benchmark data from Artificial Analysis.