← All comparisons

Llama 3 Instruct 70B vs Claude 4.1 Opus (Reasoning)

Meta vs Anthropic — side-by-side benchmark comparison

Llama 3 Instruct 70BClaude 4.1 Opus (Reasoning)
Intelligence Index8.942.0
Coding Index6.836.5
Math Index80.3
Output speed (tok/s)46.244.5
Blended price ($/1M)$1.18$32.81
Time to first token (s)0.68s8.55s
aime0.0%
aime 2580.3%
artificial analysis coding index6.8036.50
artificial analysis intelligence index8.9042.00
artificial analysis math index80.30
gpqa37.9%80.9%
hle4.4%11.9%
ifbench37.1%55.4%
lcr0.0%66.3%
livecodebench19.8%65.4%
math 50048.3%
mmlu pro57.4%88.0%
scicode18.9%40.9%
tau20.0%71.4%
terminalbench hard0.8%34.3%

Benchmark data from Artificial Analysis.