← All comparisons

Claude 4.5 Sonnet (Reasoning) vs DeepSeek V3.1 Terminus (Non-reasoning)

Anthropic vs DeepSeek — side-by-side benchmark comparison

Claude 4.5 Sonnet (Reasoning)DeepSeek V3.1 Terminus (Non-reasoning)
Intelligence Index43.028.5
Coding Index38.631.9
Math Index88.053.7
Output speed (tok/s)55.00.0
Blended price ($/1M)$6.56$0.45
Time to first token (s)7.02s0.00s
aime
aime 2588.0%53.7%
artificial analysis coding index38.6031.90
artificial analysis intelligence index43.0028.50
artificial analysis math index88.0053.70
gpqa83.4%75.1%
hle17.3%8.4%
ifbench57.3%41.2%
lcr65.7%43.3%
livecodebench71.4%52.9%
math 500
mmlu pro87.5%83.6%
scicode44.7%32.1%
tau278.1%37.1%
terminalbench hard35.6%31.8%

Benchmark data from Artificial Analysis.