← All comparisons

Claude 4.5 Sonnet (Reasoning) vs DeepSeek R1 0528 Qwen3 8B

Anthropic vs DeepSeek — side-by-side benchmark comparison

Claude 4.5 Sonnet (Reasoning)DeepSeek R1 0528 Qwen3 8B
Intelligence Index43.016.4
Coding Index38.67.8
Math Index88.063.7
Output speed (tok/s)55.00.0
Blended price ($/1M)$6.56$0.00
Time to first token (s)7.02s0.00s
aime65.0%
aime 2588.0%63.7%
artificial analysis coding index38.607.80
artificial analysis intelligence index43.0016.40
artificial analysis math index88.0063.70
gpqa83.4%61.2%
hle17.3%5.6%
ifbench57.3%19.9%
lcr65.7%13.0%
livecodebench71.4%51.3%
math 50093.2%
mmlu pro87.5%73.9%
scicode44.7%20.4%
tau278.1%0.0%
terminalbench hard35.6%1.5%

Benchmark data from Artificial Analysis.