← All comparisons

K2-V2 (high) vs Claude 4.1 Opus (Non-reasoning)

MBZUAI Institute of Foundation Models vs Anthropic — side-by-side benchmark comparison

K2-V2 (high)Claude 4.1 Opus (Non-reasoning)
Intelligence Index20.636.0
Coding Index16.1
Math Index78.3
Output speed (tok/s)0.044.7
Blended price ($/1M)$0.00$32.81
Time to first token (s)0.00s1.63s
aime
aime 2578.3%
artificial analysis coding index16.10
artificial analysis intelligence index20.6036.00
artificial analysis math index78.30
gpqa68.1%
hle9.8%
ifbench60.1%
lcr33.3%
livecodebench69.4%
math 500
mmlu pro78.6%
scicode28.6%
tau227.8%
terminalbench hard9.8%

Benchmark data from Artificial Analysis.