← All comparisons

Claude Opus 4.8 (Adaptive Reasoning, Max Effort) vs OLMo 2 32B

Anthropic vs Allen Institute for AI — side-by-side benchmark comparison

Claude Opus 4.8 (Adaptive Reasoning, Max Effort)OLMo 2 32B
Intelligence Index61.410.6
Coding Index56.72.7
Math Index3.3
Output speed (tok/s)66.90.0
Blended price ($/1M)$10.94$0.00
Time to first token (s)7.91s0.00s
aime
aime 253.3%
artificial analysis coding index56.702.70
artificial analysis intelligence index61.4010.60
artificial analysis math index3.30
gpqa92.0%32.8%
hle45.7%3.7%
ifbench62.2%38.1%
lcr67.7%0.0%
livecodebench6.8%
math 500
mmlu pro51.1%
scicode53.5%8.0%
tau294.4%0.0%
terminalbench hard58.3%0.0%

Benchmark data from Artificial Analysis.