← All comparisons

GPT-4o (May '24) vs Claude 4.1 Opus (Reasoning)

OpenAI vs Anthropic — side-by-side benchmark comparison

GPT-4o (May '24)Claude 4.1 Opus (Reasoning)
Intelligence Index14.542.0
Coding Index24.236.5
Math Index80.3
Output speed (tok/s)111.844.5
Blended price ($/1M)$7.50$32.81
Time to first token (s)0.61s8.55s
aime11.0%
aime 2580.3%
artificial analysis coding index24.2036.50
artificial analysis intelligence index14.5042.00
artificial analysis math index80.30
gpqa52.6%80.9%
hle2.8%11.9%
ifbench55.4%
lcr66.3%
livecodebench33.4%65.4%
math 50079.1%
mmlu pro74.0%88.0%
scicode30.9%40.9%
tau271.4%
terminalbench hard34.3%

Benchmark data from Artificial Analysis.