← All comparisons

gpt-oss-120b (low) vs Claude 4.5 Haiku (Non-reasoning)

OpenAI vs Anthropic — side-by-side benchmark comparison

gpt-oss-120b (low)Claude 4.5 Haiku (Non-reasoning)
Intelligence Index24.531.0
Coding Index15.529.6
Math Index66.739.0
Output speed (tok/s)370.0105.8
Blended price ($/1M)$0.26$2.19
Time to first token (s)0.49s0.65s
aime
aime 2566.7%39.0%
artificial analysis coding index15.5029.60
artificial analysis intelligence index24.5031.00
artificial analysis math index66.7039.00
gpqa67.2%64.6%
hle5.2%4.3%
ifbench58.3%42.0%
lcr43.7%43.7%
livecodebench70.7%51.1%
math 500
mmlu pro77.5%80.0%
scicode36.0%34.4%
tau245.0%32.5%
terminalbench hard5.3%27.3%

Benchmark data from Artificial Analysis.