← All comparisons

gpt-oss-120b (low) vs Claude 4.1 Opus (Non-reasoning)

OpenAI vs Anthropic — side-by-side benchmark comparison

gpt-oss-120b (low)Claude 4.1 Opus (Non-reasoning)
Intelligence Index24.536.0
Coding Index15.5
Math Index66.7
Output speed (tok/s)370.044.7
Blended price ($/1M)$0.26$32.81
Time to first token (s)0.49s1.63s
aime
aime 2566.7%
artificial analysis coding index15.50
artificial analysis intelligence index24.5036.00
artificial analysis math index66.70
gpqa67.2%
hle5.2%
ifbench58.3%
lcr43.7%
livecodebench70.7%
math 500
mmlu pro77.5%
scicode36.0%
tau245.0%
terminalbench hard5.3%

Benchmark data from Artificial Analysis.