← All comparisons

Claude 4.1 Opus (Reasoning) vs DeepSeek R1 0528 (May '25)

Anthropic vs DeepSeek — side-by-side benchmark comparison

Claude 4.1 Opus (Reasoning)DeepSeek R1 0528 (May '25)
Intelligence Index42.027.1
Coding Index36.524.0
Math Index80.376.0
Output speed (tok/s)44.50.0
Blended price ($/1M)$32.81$2.06
Time to first token (s)8.55s0.00s
aime89.3%
aime 2580.3%76.0%
artificial analysis coding index36.5024.00
artificial analysis intelligence index42.0027.10
artificial analysis math index80.3076.00
gpqa80.9%81.3%
hle11.9%14.9%
ifbench55.4%39.6%
lcr66.3%54.7%
livecodebench65.4%77.0%
math 50098.3%
mmlu pro88.0%84.9%
scicode40.9%40.3%
tau271.4%36.5%
terminalbench hard34.3%15.9%

Benchmark data from Artificial Analysis.