← All comparisons

Claude 4 Sonnet (Reasoning) vs OLMo 2 32B

Anthropic vs Allen Institute for AI — side-by-side benchmark comparison

Claude 4 Sonnet (Reasoning)OLMo 2 32B
Intelligence Index38.710.6
Coding Index34.12.7
Math Index74.33.3
Output speed (tok/s)55.50.0
Blended price ($/1M)$6.56$0.00
Time to first token (s)8.92s0.00s
aime77.3%
aime 2574.3%3.3%
artificial analysis coding index34.102.70
artificial analysis intelligence index38.7010.60
artificial analysis math index74.303.30
gpqa77.7%32.8%
hle9.6%3.7%
ifbench54.7%38.1%
lcr64.7%0.0%
livecodebench65.5%6.8%
math 50099.1%
mmlu pro84.2%51.1%
scicode40.0%8.0%
tau264.6%0.0%
terminalbench hard31.1%0.0%

Benchmark data from Artificial Analysis.