← All comparisons

Claude 3.7 Sonnet (Non-reasoning) vs OLMo 2 32B

Anthropic vs Allen Institute for AI — side-by-side benchmark comparison

Claude 3.7 Sonnet (Non-reasoning)OLMo 2 32B
Intelligence Index30.810.6
Coding Index26.72.7
Math Index21.03.3
Output speed (tok/s)0.00.0
Blended price ($/1M)$6.56$0.00
Time to first token (s)0.00s0.00s
aime22.3%
aime 2521.0%3.3%
artificial analysis coding index26.702.70
artificial analysis intelligence index30.8010.60
artificial analysis math index21.003.30
gpqa65.6%32.8%
hle4.8%3.7%
ifbench44.0%38.1%
lcr48.3%0.0%
livecodebench39.4%6.8%
math 50085.0%
mmlu pro80.3%51.1%
scicode37.6%8.0%
tau250.0%0.0%
terminalbench hard21.2%0.0%

Benchmark data from Artificial Analysis.