← All comparisons

Claude 4.5 Sonnet (Reasoning) vs DeepSeek R1 Distill Llama 70B

Anthropic vs DeepSeek — side-by-side benchmark comparison

Claude 4.5 Sonnet (Reasoning)DeepSeek R1 Distill Llama 70B
Intelligence Index43.016.0
Coding Index38.611.4
Math Index88.053.7
Output speed (tok/s)55.046.8
Blended price ($/1M)$6.56$0.79
Time to first token (s)7.02s0.33s
aime67.0%
aime 2588.0%53.7%
artificial analysis coding index38.6011.40
artificial analysis intelligence index43.0016.00
artificial analysis math index88.0053.70
gpqa83.4%40.2%
hle17.3%6.1%
ifbench57.3%27.6%
lcr65.7%11.0%
livecodebench71.4%26.6%
math 50093.5%
mmlu pro87.5%79.5%
scicode44.7%31.3%
tau278.1%21.9%
terminalbench hard35.6%1.5%

Benchmark data from Artificial Analysis.