← All comparisons

DeepSeek V4 Flash (Reasoning, High Effort) vs Claude 4.1 Opus (Reasoning)

DeepSeek vs Anthropic — side-by-side benchmark comparison

DeepSeek V4 Flash (Reasoning, High Effort)Claude 4.1 Opus (Reasoning)
Intelligence Index46.042.0
Coding Index39.836.5
Math Index80.3
Output speed (tok/s)0.044.5
Blended price ($/1M)$0.17$32.81
Time to first token (s)0.00s8.55s
aime
aime 2580.3%
artificial analysis coding index39.8036.50
artificial analysis intelligence index46.0042.00
artificial analysis math index80.30
gpqa86.7%80.9%
hle27.8%11.9%
ifbench73.5%55.4%
lcr62.7%66.3%
livecodebench65.4%
math 500
mmlu pro88.0%
scicode42.0%40.9%
tau295.6%71.4%
terminalbench hard38.6%34.3%

Benchmark data from Artificial Analysis.