← All comparisons

Claude Sonnet 4.6 (Non-reasoning, High Effort) vs OLMo 2 7B

Anthropic vs Allen Institute for AI — side-by-side benchmark comparison

Claude Sonnet 4.6 (Non-reasoning, High Effort)OLMo 2 7B
Intelligence Index44.49.3
Coding Index46.41.2
Math Index0.7
Output speed (tok/s)55.20.0
Blended price ($/1M)$6.56$0.00
Time to first token (s)1.07s0.00s
aime
aime 250.7%
artificial analysis coding index46.401.20
artificial analysis intelligence index44.409.30
artificial analysis math index70.0%
gpqa79.9%28.8%
hle13.2%5.5%
ifbench41.2%24.4%
lcr57.7%0.0%
livecodebench4.1%
math 500
mmlu pro28.2%
scicode46.9%3.7%
tau279.5%0.0%
terminalbench hard46.2%0.0%

Benchmark data from Artificial Analysis.