← All comparisons

Claude 4.1 Opus (Reasoning) vs GLM-5 (Non-reasoning)

Anthropic vs Z AI — side-by-side benchmark comparison

Claude 4.1 Opus (Reasoning)GLM-5 (Non-reasoning)
Intelligence Index42.040.6
Coding Index36.539.0
Math Index80.3
Output speed (tok/s)44.566.0
Blended price ($/1M)$32.81$1.55
Time to first token (s)8.55s1.17s
aime
aime 2580.3%
artificial analysis coding index36.5039.00
artificial analysis intelligence index42.0040.60
artificial analysis math index80.30
gpqa80.9%66.6%
hle11.9%7.2%
ifbench55.4%55.2%
lcr66.3%37.0%
livecodebench65.4%
math 500
mmlu pro88.0%
scicode40.9%38.3%
tau271.4%97.4%
terminalbench hard34.3%39.4%

Benchmark data from Artificial Analysis.