← All comparisons

Exaone 4.0 1.2B (Non-reasoning) vs Claude 3.5 Sonnet (Oct '24)

LG AI Research vs Anthropic — side-by-side benchmark comparison

Exaone 4.0 1.2B (Non-reasoning)Claude 3.5 Sonnet (Oct '24)
Intelligence Index8.115.9
Coding Index2.530.2
Math Index24.0
Output speed (tok/s)0.00.0
Blended price ($/1M)$0.00$6.56
Time to first token (s)0.00s0.00s
aime15.7%
aime 2524.0%
artificial analysis coding index2.5030.20
artificial analysis intelligence index8.1015.90
artificial analysis math index24.00
gpqa42.4%59.9%
hle5.8%3.9%
ifbench25.3%
lcr0.0%
livecodebench29.3%38.1%
math 50077.1%
mmlu pro50.0%77.2%
scicode7.4%36.6%
tau220.5%
terminalbench hard0.0%

Benchmark data from Artificial Analysis.