← All comparisons

Llama 3.1 Instruct 8B vs Claude 4.1 Opus (Reasoning)

Meta vs Anthropic — side-by-side benchmark comparison

Llama 3.1 Instruct 8BClaude 4.1 Opus (Reasoning)
Intelligence Index11.842.0
Coding Index4.936.5
Math Index4.380.3
Output speed (tok/s)212.744.5
Blended price ($/1M)$0.10$32.81
Time to first token (s)0.45s8.55s
aime7.7%
aime 254.3%80.3%
artificial analysis coding index4.9036.50
artificial analysis intelligence index11.8042.00
artificial analysis math index4.3080.30
gpqa25.9%80.9%
hle5.1%11.9%
ifbench28.6%55.4%
lcr15.7%66.3%
livecodebench11.6%65.4%
math 50051.9%
mmlu pro47.6%88.0%
scicode13.2%40.9%
tau216.4%71.4%
terminalbench hard0.8%34.3%

Benchmark data from Artificial Analysis.