← All comparisons

GPT-4o (Nov '24) vs Claude 3.5 Sonnet (June '24)

OpenAI vs Anthropic — side-by-side benchmark comparison

GPT-4o (Nov '24)Claude 3.5 Sonnet (June '24)
Intelligence Index17.314.2
Coding Index16.726.0
Math Index6.0
Output speed (tok/s)158.50.0
Blended price ($/1M)$4.38$6.56
Time to first token (s)0.54s0.00s
aime15.0%9.7%
aime 256.0%
artificial analysis coding index16.7026.00
artificial analysis intelligence index17.3014.20
artificial analysis math index6.00
gpqa54.3%56.0%
hle3.3%3.7%
ifbench34.3%
lcr0.0%
livecodebench30.9%
math 50075.9%69.5%
mmlu pro74.8%75.1%
scicode33.3%31.6%
tau225.1%
terminalbench hard8.3%

Benchmark data from Artificial Analysis.