← All comparisons

Claude 4.1 Opus (Reasoning) vs OLMo 2 7B

Anthropic vs Allen Institute for AI — side-by-side benchmark comparison

Claude 4.1 Opus (Reasoning)OLMo 2 7B
Intelligence Index42.09.3
Coding Index36.51.2
Math Index80.30.7
Output speed (tok/s)44.50.0
Blended price ($/1M)$32.81$0.00
Time to first token (s)8.55s0.00s
aime
aime 2580.3%0.7%
artificial analysis coding index36.501.20
artificial analysis intelligence index42.009.30
artificial analysis math index80.3070.0%
gpqa80.9%28.8%
hle11.9%5.5%
ifbench55.4%24.4%
lcr66.3%0.0%
livecodebench65.4%4.1%
math 500
mmlu pro88.0%28.2%
scicode40.9%3.7%
tau271.4%0.0%
terminalbench hard34.3%0.0%

Benchmark data from Artificial Analysis.