← All comparisons

Claude 4.1 Opus (Reasoning) vs GLM-4.5-Air

Anthropic vs Z AI — side-by-side benchmark comparison

Claude 4.1 Opus (Reasoning)GLM-4.5-Air
Intelligence Index42.023.2
Coding Index36.523.8
Math Index80.380.7
Output speed (tok/s)44.586.9
Blended price ($/1M)$32.81$0.37
Time to first token (s)8.55s1.72s
aime67.3%
aime 2580.3%80.7%
artificial analysis coding index36.5023.80
artificial analysis intelligence index42.0023.20
artificial analysis math index80.3080.70
gpqa80.9%73.3%
hle11.9%6.8%
ifbench55.4%37.6%
lcr66.3%43.7%
livecodebench65.4%68.4%
math 50096.5%
mmlu pro88.0%81.5%
scicode40.9%30.6%
tau271.4%46.5%
terminalbench hard34.3%20.5%

Benchmark data from Artificial Analysis.