← All comparisons

Llama 3.3 Instruct 70B vs Claude 3.5 Sonnet (June '24)

Meta vs Anthropic — side-by-side benchmark comparison

Llama 3.3 Instruct 70BClaude 3.5 Sonnet (June '24)
Intelligence Index14.514.2
Coding Index10.726.0
Math Index7.7
Output speed (tok/s)88.10.0
Blended price ($/1M)$0.62$6.56
Time to first token (s)0.59s0.00s
aime30.0%9.7%
aime 257.7%
artificial analysis coding index10.7026.00
artificial analysis intelligence index14.5014.20
artificial analysis math index7.70
gpqa49.8%56.0%
hle4.0%3.7%
ifbench47.1%
lcr15.0%
livecodebench28.8%
math 50077.3%69.5%
mmlu pro71.3%75.1%
scicode26.0%31.6%
tau226.6%
terminalbench hard3.0%

Benchmark data from Artificial Analysis.