← All comparisons

GPT-4o (March 2025, chatgpt-4o-latest) vs Claude 4.1 Opus (Reasoning)

OpenAI vs Anthropic — side-by-side benchmark comparison

GPT-4o (March 2025, chatgpt-4o-latest)Claude 4.1 Opus (Reasoning)
Intelligence Index18.642.0
Coding Index36.5
Math Index25.780.3
Output speed (tok/s)0.044.5
Blended price ($/1M)$0.00$32.81
Time to first token (s)0.00s8.55s
aime32.7%
aime 2525.7%80.3%
artificial analysis coding index36.50
artificial analysis intelligence index18.6042.00
artificial analysis math index25.7080.30
gpqa65.5%80.9%
hle5.0%11.9%
ifbench55.4%
lcr66.3%
livecodebench42.5%65.4%
math 50089.3%
mmlu pro80.3%88.0%
scicode36.6%40.9%
tau271.4%
terminalbench hard34.3%

Benchmark data from Artificial Analysis.