← All comparisons

Claude 4.5 Sonnet (Reasoning) vs DeepSeek V3.1 Terminus (Reasoning)

Anthropic vs DeepSeek — side-by-side benchmark comparison

Claude 4.5 Sonnet (Reasoning)DeepSeek V3.1 Terminus (Reasoning)
Intelligence Index43.033.9
Coding Index38.633.7
Math Index88.089.7
Output speed (tok/s)55.00.0
Blended price ($/1M)$6.56$1.91
Time to first token (s)7.02s0.00s
aime
aime 2588.0%89.7%
artificial analysis coding index38.6033.70
artificial analysis intelligence index43.0033.90
artificial analysis math index88.0089.70
gpqa83.4%79.2%
hle17.3%15.2%
ifbench57.3%57.0%
lcr65.7%65.0%
livecodebench71.4%79.8%
math 500
mmlu pro87.5%85.1%
scicode44.7%40.6%
tau278.1%37.1%
terminalbench hard35.6%30.3%

Benchmark data from Artificial Analysis.