← All comparisons

Llama 3.1 Instruct 70B vs Claude 4.1 Opus (Reasoning)

Meta vs Anthropic — side-by-side benchmark comparison

Llama 3.1 Instruct 70BClaude 4.1 Opus (Reasoning)
Intelligence Index12.542.0
Coding Index10.936.5
Math Index4.080.3
Output speed (tok/s)35.344.5
Blended price ($/1M)$0.56$32.81
Time to first token (s)0.54s8.55s
aime17.3%
aime 254.0%80.3%
artificial analysis coding index10.9036.50
artificial analysis intelligence index12.5042.00
artificial analysis math index4.0080.30
gpqa40.9%80.9%
hle4.6%11.9%
ifbench34.4%55.4%
lcr6.3%66.3%
livecodebench23.2%65.4%
math 50064.9%
mmlu pro67.6%88.0%
scicode26.7%40.9%
tau215.2%71.4%
terminalbench hard3.0%34.3%

Benchmark data from Artificial Analysis.