← All comparisons

Claude 4.1 Opus (Reasoning) vs Claude 4 Opus (Reasoning)

Anthropic vs Anthropic — side-by-side benchmark comparison

Claude 4.1 Opus (Reasoning)Claude 4 Opus (Reasoning)
Intelligence Index42.039.0
Coding Index36.534.0
Math Index80.373.3
Output speed (tok/s)44.542.1
Blended price ($/1M)$32.81$32.81
Time to first token (s)8.55s6.80s
aime75.7%
aime 2580.3%73.3%
artificial analysis coding index36.5034.00
artificial analysis intelligence index42.0039.00
artificial analysis math index80.3073.30
gpqa80.9%79.6%
hle11.9%11.7%
ifbench55.4%53.7%
lcr66.3%33.7%
livecodebench65.4%63.6%
math 50098.2%
mmlu pro88.0%87.3%
scicode40.9%39.8%
tau271.4%73.4%
terminalbench hard34.3%31.1%

Benchmark data from Artificial Analysis.