← All comparisons

GLM-5.1 (Reasoning) vs Claude 4.1 Opus (Non-reasoning)

Z AI vs Anthropic — side-by-side benchmark comparison

GLM-5.1 (Reasoning)Claude 4.1 Opus (Non-reasoning)
Intelligence Index51.436.0
Coding Index43.4
Math Index
Output speed (tok/s)61.244.7
Blended price ($/1M)$2.15$32.81
Time to first token (s)0.86s1.63s
aime
aime 25
artificial analysis coding index43.40
artificial analysis intelligence index51.4036.00
artificial analysis math index
gpqa86.8%
hle28.0%
ifbench76.3%
lcr62.3%
livecodebench
math 500
mmlu pro
scicode43.8%
tau297.7%
terminalbench hard43.2%

Benchmark data from Artificial Analysis.