← All comparisons

Claude Sonnet 4.6 (Adaptive Reasoning, Max Effort) vs OLMo 2 7B

Anthropic vs Allen Institute for AI — side-by-side benchmark comparison

Claude Sonnet 4.6 (Adaptive Reasoning, Max Effort)OLMo 2 7B
Intelligence Index51.79.3
Coding Index50.91.2
Math Index0.7
Output speed (tok/s)68.20.0
Blended price ($/1M)$6.56$0.00
Time to first token (s)55.35s0.00s
aime
aime 250.7%
artificial analysis coding index50.901.20
artificial analysis intelligence index51.709.30
artificial analysis math index70.0%
gpqa87.5%28.8%
hle30.0%5.5%
ifbench56.6%24.4%
lcr70.7%0.0%
livecodebench4.1%
math 500
mmlu pro28.2%
scicode46.8%3.7%
tau275.7%0.0%
terminalbench hard53.0%0.0%

Benchmark data from Artificial Analysis.