← All comparisons

Claude 4.5 Sonnet (Reasoning) vs DeepSeek V3 (Dec '24)

Anthropic vs DeepSeek — side-by-side benchmark comparison

Claude 4.5 Sonnet (Reasoning)DeepSeek V3 (Dec '24)
Intelligence Index43.016.5
Coding Index38.616.4
Math Index88.026.0
Output speed (tok/s)55.00.0
Blended price ($/1M)$6.56$0.52
Time to first token (s)7.02s0.00s
aime25.3%
aime 2588.0%26.0%
artificial analysis coding index38.6016.40
artificial analysis intelligence index43.0016.50
artificial analysis math index88.0026.00
gpqa83.4%55.7%
hle17.3%3.6%
ifbench57.3%34.8%
lcr65.7%29.0%
livecodebench71.4%35.9%
math 50088.7%
mmlu pro87.5%75.2%
scicode44.7%35.4%
tau278.1%22.8%
terminalbench hard35.6%6.8%

Benchmark data from Artificial Analysis.