← All comparisons

Claude 3.5 Sonnet (Oct '24) vs DeepSeek R1 (Jan '25)

Anthropic vs DeepSeek — side-by-side benchmark comparison

Claude 3.5 Sonnet (Oct '24)DeepSeek R1 (Jan '25)
Intelligence Index15.918.8
Coding Index30.215.9
Math Index68.0
Output speed (tok/s)0.00.0
Blended price ($/1M)$6.56$2.43
Time to first token (s)0.00s0.00s
aime15.7%68.3%
aime 2568.0%
artificial analysis coding index30.2015.90
artificial analysis intelligence index15.9018.80
artificial analysis math index68.00
gpqa59.9%70.8%
hle3.9%9.3%
ifbench39.0%
lcr52.3%
livecodebench38.1%61.7%
math 50077.1%96.6%
mmlu pro77.2%84.4%
scicode36.6%35.7%
tau211.4%
terminalbench hard6.1%

Benchmark data from Artificial Analysis.