← All comparisons

Claude 4.5 Sonnet (Reasoning) vs DeepSeek V3.1 (Reasoning)

Anthropic vs DeepSeek — side-by-side benchmark comparison

Claude 4.5 Sonnet (Reasoning)DeepSeek V3.1 (Reasoning)
Intelligence Index43.027.7
Coding Index38.629.7
Math Index88.089.7
Output speed (tok/s)55.00.0
Blended price ($/1M)$6.56$0.86
Time to first token (s)7.02s0.00s
aime
aime 2588.0%89.7%
artificial analysis coding index38.6029.70
artificial analysis intelligence index43.0027.70
artificial analysis math index88.0089.70
gpqa83.4%77.9%
hle17.3%13.0%
ifbench57.3%41.5%
lcr65.7%53.3%
livecodebench71.4%78.4%
math 500
mmlu pro87.5%85.1%
scicode44.7%39.1%
tau278.1%37.4%
terminalbench hard35.6%25.0%

Benchmark data from Artificial Analysis.