← All comparisons

Claude 4.5 Sonnet (Reasoning) vs OLMo 2 7B

Anthropic vs Allen Institute for AI — side-by-side benchmark comparison

Claude 4.5 Sonnet (Reasoning)OLMo 2 7B
Intelligence Index43.09.3
Coding Index38.61.2
Math Index88.00.7
Output speed (tok/s)55.00.0
Blended price ($/1M)$6.56$0.00
Time to first token (s)7.02s0.00s
aime
aime 2588.0%0.7%
artificial analysis coding index38.601.20
artificial analysis intelligence index43.009.30
artificial analysis math index88.0070.0%
gpqa83.4%28.8%
hle17.3%5.5%
ifbench57.3%24.4%
lcr65.7%0.0%
livecodebench71.4%4.1%
math 500
mmlu pro87.5%28.2%
scicode44.7%3.7%
tau278.1%0.0%
terminalbench hard35.6%0.0%

Benchmark data from Artificial Analysis.