← All comparisons

K-EXAONE (Reasoning) vs Claude 3.5 Sonnet (June '24)

LG AI Research vs Anthropic — side-by-side benchmark comparison

K-EXAONE (Reasoning)Claude 3.5 Sonnet (June '24)
Intelligence Index32.114.2
Coding Index27.026.0
Math Index90.3
Output speed (tok/s)0.00.0
Blended price ($/1M)$0.00$6.56
Time to first token (s)0.00s0.00s
aime9.7%
aime 2590.3%
artificial analysis coding index27.0026.00
artificial analysis intelligence index32.1014.20
artificial analysis math index90.30
gpqa78.3%56.0%
hle13.1%3.7%
ifbench64.7%
lcr55.7%
livecodebench76.8%
math 50069.5%
mmlu pro83.8%75.1%
scicode35.6%31.6%
tau274.3%
terminalbench hard22.7%

Benchmark data from Artificial Analysis.