← All comparisons

gpt-oss-20B (low) vs Claude 4.1 Opus (Reasoning)

OpenAI vs Anthropic — side-by-side benchmark comparison

gpt-oss-20B (low)Claude 4.1 Opus (Reasoning)
Intelligence Index20.842.0
Coding Index14.436.5
Math Index62.380.3
Output speed (tok/s)273.044.5
Blended price ($/1M)$0.10$32.81
Time to first token (s)0.50s8.55s
aime
aime 2562.3%80.3%
artificial analysis coding index14.4036.50
artificial analysis intelligence index20.8042.00
artificial analysis math index62.3080.30
gpqa61.1%80.9%
hle5.1%11.9%
ifbench57.8%55.4%
lcr31.0%66.3%
livecodebench65.2%65.4%
math 500
mmlu pro71.8%88.0%
scicode34.0%40.9%
tau250.3%71.4%
terminalbench hard4.5%34.3%

Benchmark data from Artificial Analysis.