← All comparisons

gpt-oss-20B (low) vs Claude 4.1 Opus (Non-reasoning)

OpenAI vs Anthropic — side-by-side benchmark comparison

gpt-oss-20B (low)Claude 4.1 Opus (Non-reasoning)
Intelligence Index20.836.0
Coding Index14.4
Math Index62.3
Output speed (tok/s)273.044.7
Blended price ($/1M)$0.10$32.81
Time to first token (s)0.50s1.63s
aime
aime 2562.3%
artificial analysis coding index14.40
artificial analysis intelligence index20.8036.00
artificial analysis math index62.30
gpqa61.1%
hle5.1%
ifbench57.8%
lcr31.0%
livecodebench65.2%
math 500
mmlu pro71.8%
scicode34.0%
tau250.3%
terminalbench hard4.5%

Benchmark data from Artificial Analysis.