← All comparisons

Claude 4.1 Opus (Non-reasoning) vs DeepSeek R1 (Jan '25)

Anthropic vs DeepSeek — side-by-side benchmark comparison

Claude 4.1 Opus (Non-reasoning)DeepSeek R1 (Jan '25)
Intelligence Index36.018.8
Coding Index15.9
Math Index68.0
Output speed (tok/s)44.70.0
Blended price ($/1M)$32.81$2.43
Time to first token (s)1.63s0.00s
aime68.3%
aime 2568.0%
artificial analysis coding index15.90
artificial analysis intelligence index36.0018.80
artificial analysis math index68.00
gpqa70.8%
hle9.3%
ifbench39.0%
lcr52.3%
livecodebench61.7%
math 50096.6%
mmlu pro84.4%
scicode35.7%
tau211.4%
terminalbench hard6.1%

Benchmark data from Artificial Analysis.