← All comparisons

K-EXAONE (Reasoning) vs Claude 3.5 Sonnet (Oct '24)

LG AI Research vs Anthropic — side-by-side benchmark comparison

K-EXAONE (Reasoning)Claude 3.5 Sonnet (Oct '24)
Intelligence Index32.115.9
Coding Index27.030.2
Math Index90.3
Output speed (tok/s)0.00.0
Blended price ($/1M)$0.00$6.56
Time to first token (s)0.00s0.00s
aime15.7%
aime 2590.3%
artificial analysis coding index27.0030.20
artificial analysis intelligence index32.1015.90
artificial analysis math index90.30
gpqa78.3%59.9%
hle13.1%3.9%
ifbench64.7%
lcr55.7%
livecodebench76.8%38.1%
math 50077.1%
mmlu pro83.8%77.2%
scicode35.6%36.6%
tau274.3%
terminalbench hard22.7%

Benchmark data from Artificial Analysis.