← All comparisons

Exaone 4.0 1.2B (Non-reasoning) vs Claude 3.5 Sonnet (June '24)

LG AI Research vs Anthropic — side-by-side benchmark comparison

Exaone 4.0 1.2B (Non-reasoning)Claude 3.5 Sonnet (June '24)
Intelligence Index8.114.2
Coding Index2.526.0
Math Index24.0
Output speed (tok/s)0.00.0
Blended price ($/1M)$0.00$6.56
Time to first token (s)0.00s0.00s
aime9.7%
aime 2524.0%
artificial analysis coding index2.5026.00
artificial analysis intelligence index8.1014.20
artificial analysis math index24.00
gpqa42.4%56.0%
hle5.8%3.7%
ifbench25.3%
lcr0.0%
livecodebench29.3%
math 50069.5%
mmlu pro50.0%75.1%
scicode7.4%31.6%
tau220.5%
terminalbench hard0.0%

Benchmark data from Artificial Analysis.