← All comparisons

GPT-5.1 (Non-reasoning) vs Claude 3.5 Sonnet (June '24)

OpenAI vs Anthropic — side-by-side benchmark comparison

GPT-5.1 (Non-reasoning)Claude 3.5 Sonnet (June '24)
Intelligence Index27.414.2
Coding Index27.326.0
Math Index38.0
Output speed (tok/s)129.50.0
Blended price ($/1M)$3.44$6.56
Time to first token (s)0.72s0.00s
aime9.7%
aime 2538.0%
artificial analysis coding index27.3026.00
artificial analysis intelligence index27.4014.20
artificial analysis math index38.00
gpqa64.3%56.0%
hle5.2%3.7%
ifbench43.2%
lcr44.0%
livecodebench49.4%
math 50069.5%
mmlu pro80.1%75.1%
scicode36.5%31.6%
tau246.5%
terminalbench hard22.7%

Benchmark data from Artificial Analysis.