← All comparisons

Claude 4.1 Opus (Reasoning) vs Kimi K2 Thinking

Anthropic vs Kimi — side-by-side benchmark comparison

Claude 4.1 Opus (Reasoning)Kimi K2 Thinking
Intelligence Index42.040.9
Coding Index36.534.8
Math Index80.394.7
Output speed (tok/s)44.5130.3
Blended price ($/1M)$32.81$1.07
Time to first token (s)8.55s0.84s
aime
aime 2580.3%94.7%
artificial analysis coding index36.5034.80
artificial analysis intelligence index42.0040.90
artificial analysis math index80.3094.70
gpqa80.9%83.8%
hle11.9%22.3%
ifbench55.4%68.1%
lcr66.3%66.3%
livecodebench65.4%85.3%
math 500
mmlu pro88.0%84.8%
scicode40.9%42.4%
tau271.4%93.0%
terminalbench hard34.3%31.1%

Benchmark data from Artificial Analysis.