← All comparisons

GPT-3.5 Turbo vs Claude 4.1 Opus (Reasoning)

OpenAI vs Anthropic — side-by-side benchmark comparison

GPT-3.5 TurboClaude 4.1 Opus (Reasoning)
Intelligence Index9.042.0
Coding Index10.736.5
Math Index80.3
Output speed (tok/s)116.944.5
Blended price ($/1M)$0.75$32.81
Time to first token (s)0.56s8.55s
aime
aime 2580.3%
artificial analysis coding index10.7036.50
artificial analysis intelligence index9.0042.00
artificial analysis math index80.30
gpqa29.7%80.9%
hle11.9%
ifbench55.4%
lcr66.3%
livecodebench65.4%
math 50044.1%
mmlu pro46.2%88.0%
scicode40.9%
tau271.4%
terminalbench hard34.3%

Benchmark data from Artificial Analysis.