← All comparisons

GPT-5.1 (Non-reasoning) vs Claude 3.5 Sonnet (Oct '24)

OpenAI vs Anthropic — side-by-side benchmark comparison

GPT-5.1 (Non-reasoning)Claude 3.5 Sonnet (Oct '24)
Intelligence Index27.415.9
Coding Index27.330.2
Math Index38.0
Output speed (tok/s)129.50.0
Blended price ($/1M)$3.44$6.56
Time to first token (s)0.72s0.00s
aime15.7%
aime 2538.0%
artificial analysis coding index27.3030.20
artificial analysis intelligence index27.4015.90
artificial analysis math index38.00
gpqa64.3%59.9%
hle5.2%3.9%
ifbench43.2%
lcr44.0%
livecodebench49.4%38.1%
math 50077.1%
mmlu pro80.1%77.2%
scicode36.5%36.6%
tau246.5%
terminalbench hard22.7%

Benchmark data from Artificial Analysis.