← All comparisons

gpt-oss-20B (low) vs Claude 3.5 Sonnet (Oct '24)

OpenAI vs Anthropic — side-by-side benchmark comparison

gpt-oss-20B (low)Claude 3.5 Sonnet (Oct '24)
Intelligence Index20.815.9
Coding Index14.430.2
Math Index62.3
Output speed (tok/s)273.00.0
Blended price ($/1M)$0.10$6.56
Time to first token (s)0.50s0.00s
aime15.7%
aime 2562.3%
artificial analysis coding index14.4030.20
artificial analysis intelligence index20.8015.90
artificial analysis math index62.30
gpqa61.1%59.9%
hle5.1%3.9%
ifbench57.8%
lcr31.0%
livecodebench65.2%38.1%
math 50077.1%
mmlu pro71.8%77.2%
scicode34.0%36.6%
tau250.3%
terminalbench hard4.5%

Benchmark data from Artificial Analysis.