← All comparisons

Claude 4.5 Sonnet (Reasoning) vs DeepSeek V3.1 (Non-reasoning)

Anthropic vs DeepSeek — side-by-side benchmark comparison

Claude 4.5 Sonnet (Reasoning)DeepSeek V3.1 (Non-reasoning)
Intelligence Index43.028.1
Coding Index38.628.4
Math Index88.049.7
Output speed (tok/s)55.00.0
Blended price ($/1M)$6.56$0.83
Time to first token (s)7.02s0.00s
aime
aime 2588.0%49.7%
artificial analysis coding index38.6028.40
artificial analysis intelligence index43.0028.10
artificial analysis math index88.0049.70
gpqa83.4%73.5%
hle17.3%6.3%
ifbench57.3%37.8%
lcr65.7%45.0%
livecodebench71.4%57.7%
math 500
mmlu pro87.5%83.3%
scicode44.7%36.7%
tau278.1%34.8%
terminalbench hard35.6%24.2%

Benchmark data from Artificial Analysis.