← All comparisons

Claude 3.5 Sonnet (June '24) vs Kimi K2 Thinking

Anthropic vs Kimi — side-by-side benchmark comparison

Claude 3.5 Sonnet (June '24)Kimi K2 Thinking
Intelligence Index14.240.9
Coding Index26.034.8
Math Index94.7
Output speed (tok/s)0.0130.3
Blended price ($/1M)$6.56$1.07
Time to first token (s)0.00s0.84s
aime9.7%
aime 2594.7%
artificial analysis coding index26.0034.80
artificial analysis intelligence index14.2040.90
artificial analysis math index94.70
gpqa56.0%83.8%
hle3.7%22.3%
ifbench68.1%
lcr66.3%
livecodebench85.3%
math 50069.5%
mmlu pro75.1%84.8%
scicode31.6%42.4%
tau293.0%
terminalbench hard31.1%

Benchmark data from Artificial Analysis.