← All comparisons

Claude 4.1 Opus (Reasoning) vs Claude 4 Opus (Non-reasoning)

Anthropic vs Anthropic — side-by-side benchmark comparison

Claude 4.1 Opus (Reasoning)Claude 4 Opus (Non-reasoning)
Intelligence Index42.033.0
Coding Index36.5
Math Index80.336.3
Output speed (tok/s)44.543.3
Blended price ($/1M)$32.81$32.81
Time to first token (s)8.55s1.66s
aime56.3%
aime 2580.3%36.3%
artificial analysis coding index36.50
artificial analysis intelligence index42.0033.00
artificial analysis math index80.3036.30
gpqa80.9%70.1%
hle11.9%5.9%
ifbench55.4%43.3%
lcr66.3%36.0%
livecodebench65.4%54.2%
math 50094.1%
mmlu pro88.0%86.0%
scicode40.9%40.9%
tau271.4%
terminalbench hard34.3%

Benchmark data from Artificial Analysis.