← All comparisons

Claude 4.1 Opus (Reasoning) vs DeepSeek V3.1 (Reasoning)

Anthropic vs DeepSeek — side-by-side benchmark comparison

Claude 4.1 Opus (Reasoning)DeepSeek V3.1 (Reasoning)
Intelligence Index42.027.7
Coding Index36.529.7
Math Index80.389.7
Output speed (tok/s)44.50.0
Blended price ($/1M)$32.81$0.86
Time to first token (s)8.55s0.00s
aime
aime 2580.3%89.7%
artificial analysis coding index36.5029.70
artificial analysis intelligence index42.0027.70
artificial analysis math index80.3089.70
gpqa80.9%77.9%
hle11.9%13.0%
ifbench55.4%41.5%
lcr66.3%53.3%
livecodebench65.4%78.4%
math 500
mmlu pro88.0%85.1%
scicode40.9%39.1%
tau271.4%37.4%
terminalbench hard34.3%25.0%

Benchmark data from Artificial Analysis.