← All comparisons

K2 Think V2 vs Claude 4.1 Opus (Non-reasoning)

MBZUAI Institute of Foundation Models vs Anthropic — side-by-side benchmark comparison

K2 Think V2Claude 4.1 Opus (Non-reasoning)
Intelligence Index24.136.0
Coding Index15.5
Math Index
Output speed (tok/s)0.044.7
Blended price ($/1M)$0.00$32.81
Time to first token (s)0.00s1.63s
aime
aime 25
artificial analysis coding index15.50
artificial analysis intelligence index24.1036.00
artificial analysis math index
gpqa71.3%
hle9.5%
ifbench62.8%
lcr52.7%
livecodebench
math 500
mmlu pro
scicode33.0%
tau225.4%
terminalbench hard6.8%

Benchmark data from Artificial Analysis.