← All comparisons

Claude 4.1 Opus (Reasoning) vs DeepSeek V3.1 Terminus (Non-reasoning)

Anthropic vs DeepSeek — side-by-side benchmark comparison

Claude 4.1 Opus (Reasoning)DeepSeek V3.1 Terminus (Non-reasoning)
Intelligence Index42.028.5
Coding Index36.531.9
Math Index80.353.7
Output speed (tok/s)44.50.0
Blended price ($/1M)$32.81$0.45
Time to first token (s)8.55s0.00s
aime
aime 2580.3%53.7%
artificial analysis coding index36.5031.90
artificial analysis intelligence index42.0028.50
artificial analysis math index80.3053.70
gpqa80.9%75.1%
hle11.9%8.4%
ifbench55.4%41.2%
lcr66.3%43.3%
livecodebench65.4%52.9%
math 500
mmlu pro88.0%83.6%
scicode40.9%32.1%
tau271.4%37.1%
terminalbench hard34.3%31.8%

Benchmark data from Artificial Analysis.