← All comparisons

Llama 4 Scout vs Claude 4.1 Opus (Reasoning)

Meta vs Anthropic — side-by-side benchmark comparison

Llama 4 ScoutClaude 4.1 Opus (Reasoning)
Intelligence Index13.542.0
Coding Index6.736.5
Math Index14.080.3
Output speed (tok/s)113.844.5
Blended price ($/1M)$0.29$32.81
Time to first token (s)0.56s8.55s
aime28.3%
aime 2514.0%80.3%
artificial analysis coding index6.7036.50
artificial analysis intelligence index13.5042.00
artificial analysis math index14.0080.30
gpqa58.7%80.9%
hle4.3%11.9%
ifbench39.5%55.4%
lcr25.8%66.3%
livecodebench29.9%65.4%
math 50084.4%
mmlu pro75.2%88.0%
scicode17.0%40.9%
tau215.5%71.4%
terminalbench hard1.5%34.3%

Benchmark data from Artificial Analysis.