← All comparisons

Claude 4.1 Opus (Reasoning) vs DeepSeek V3.2 Exp (Reasoning)

Anthropic vs DeepSeek — side-by-side benchmark comparison

Claude 4.1 Opus (Reasoning)DeepSeek V3.2 Exp (Reasoning)
Intelligence Index42.032.9
Coding Index36.533.3
Math Index80.387.7
Output speed (tok/s)44.50.0
Blended price ($/1M)$32.81$0.31
Time to first token (s)8.55s0.00s
aime
aime 2580.3%87.7%
artificial analysis coding index36.5033.30
artificial analysis intelligence index42.0032.90
artificial analysis math index80.3087.70
gpqa80.9%79.7%
hle11.9%13.8%
ifbench55.4%54.1%
lcr66.3%69.0%
livecodebench65.4%78.9%
math 500
mmlu pro88.0%85.0%
scicode40.9%37.7%
tau271.4%33.9%
terminalbench hard34.3%31.1%

Benchmark data from Artificial Analysis.