← All comparisons

Claude 4.1 Opus (Reasoning) vs DeepSeek R1 Distill Qwen 32B

Anthropic vs DeepSeek — side-by-side benchmark comparison

Claude 4.1 Opus (Reasoning)DeepSeek R1 Distill Qwen 32B
Intelligence Index42.017.2
Coding Index36.5
Math Index80.363.0
Output speed (tok/s)44.50.0
Blended price ($/1M)$32.81$0.00
Time to first token (s)8.55s0.00s
aime68.7%
aime 2580.3%63.0%
artificial analysis coding index36.50
artificial analysis intelligence index42.0017.20
artificial analysis math index80.3063.00
gpqa80.9%61.5%
hle11.9%5.5%
ifbench55.4%22.9%
lcr66.3%9.7%
livecodebench65.4%27.0%
math 50094.1%
mmlu pro88.0%73.9%
scicode40.9%37.6%
tau271.4%
terminalbench hard34.3%

Benchmark data from Artificial Analysis.