← All comparisons

GPT-4.1 nano vs Claude 3.5 Sonnet (June '24)

OpenAI vs Anthropic — side-by-side benchmark comparison

GPT-4.1 nanoClaude 3.5 Sonnet (June '24)
Intelligence Index13.014.2
Coding Index11.226.0
Math Index24.0
Output speed (tok/s)178.90.0
Blended price ($/1M)$0.17$6.56
Time to first token (s)0.40s0.00s
aime23.7%9.7%
aime 2524.0%
artificial analysis coding index11.2026.00
artificial analysis intelligence index13.0014.20
artificial analysis math index24.00
gpqa51.2%56.0%
hle3.9%3.7%
ifbench32.0%
lcr17.0%
livecodebench32.6%
math 50084.8%69.5%
mmlu pro65.7%75.1%
scicode25.9%31.6%
tau217.3%
terminalbench hard3.8%

Benchmark data from Artificial Analysis.