← All comparisons

Llama 3.2 Instruct 1B vs Claude 4.1 Opus (Reasoning)

Meta vs Anthropic — side-by-side benchmark comparison

Llama 3.2 Instruct 1BClaude 4.1 Opus (Reasoning)
Intelligence Index6.342.0
Coding Index0.636.5
Math Index0.080.3
Output speed (tok/s)89.744.5
Blended price ($/1M)$0.05$32.81
Time to first token (s)0.59s8.55s
aime0.0%
aime 250.0%80.3%
artificial analysis coding index60.0%36.50
artificial analysis intelligence index6.3042.00
artificial analysis math index0.0%80.30
gpqa19.6%80.9%
hle5.3%11.9%
ifbench22.8%55.4%
lcr5.0%66.3%
livecodebench1.9%65.4%
math 50014.0%
mmlu pro20.0%88.0%
scicode1.7%40.9%
tau20.0%71.4%
terminalbench hard0.0%34.3%

Benchmark data from Artificial Analysis.