← All comparisons

Claude 4.1 Opus (Reasoning) vs DeepSeek R1 Distill Llama 70B

Anthropic vs DeepSeek — side-by-side benchmark comparison

Claude 4.1 Opus (Reasoning)DeepSeek R1 Distill Llama 70B
Intelligence Index42.016.0
Coding Index36.511.4
Math Index80.353.7
Output speed (tok/s)44.546.8
Blended price ($/1M)$32.81$0.79
Time to first token (s)8.55s0.33s
aime67.0%
aime 2580.3%53.7%
artificial analysis coding index36.5011.40
artificial analysis intelligence index42.0016.00
artificial analysis math index80.3053.70
gpqa80.9%40.2%
hle11.9%6.1%
ifbench55.4%27.6%
lcr66.3%11.0%
livecodebench65.4%26.6%
math 50093.5%
mmlu pro88.0%79.5%
scicode40.9%31.3%
tau271.4%21.9%
terminalbench hard34.3%1.5%

Benchmark data from Artificial Analysis.