← All comparisons

gpt-oss-120b (high) vs Claude 4.5 Haiku (Reasoning)

OpenAI vs Anthropic — side-by-side benchmark comparison

gpt-oss-120b (high)Claude 4.5 Haiku (Reasoning)
Intelligence Index33.337.1
Coding Index28.632.6
Math Index93.483.7
Output speed (tok/s)356.8142.2
Blended price ($/1M)$0.26$2.19
Time to first token (s)0.51s10.48s
aime
aime 2593.4%83.7%
artificial analysis coding index28.6032.60
artificial analysis intelligence index33.3037.10
artificial analysis math index93.4083.70
gpqa78.2%67.2%
hle18.5%9.7%
ifbench69.0%54.3%
lcr50.7%70.3%
livecodebench87.8%61.5%
math 500
mmlu pro80.8%76.0%
scicode38.9%43.3%
tau265.8%54.7%
terminalbench hard23.5%27.3%

Benchmark data from Artificial Analysis.