← All comparisons

EXAONE 4.0 32B (Non-reasoning) vs Claude 3.5 Sonnet (Oct '24)

LG AI Research vs Anthropic — side-by-side benchmark comparison

EXAONE 4.0 32B (Non-reasoning)Claude 3.5 Sonnet (Oct '24)
Intelligence Index11.715.9
Coding Index9.430.2
Math Index39.3
Output speed (tok/s)0.00.0
Blended price ($/1M)$0.00$6.56
Time to first token (s)0.00s0.00s
aime47.0%15.7%
aime 2539.3%
artificial analysis coding index9.4030.20
artificial analysis intelligence index11.7015.90
artificial analysis math index39.30
gpqa62.8%59.9%
hle4.9%3.9%
ifbench33.5%
lcr8.0%
livecodebench47.2%38.1%
math 50093.9%77.1%
mmlu pro76.8%77.2%
scicode25.2%36.6%
tau24.1%
terminalbench hard1.5%

Benchmark data from Artificial Analysis.