← All comparisons

Claude 4.1 Opus (Reasoning) vs GLM-4.6 (Reasoning)

Anthropic vs Z AI — side-by-side benchmark comparison

Claude 4.1 Opus (Reasoning)GLM-4.6 (Reasoning)
Intelligence Index42.032.5
Coding Index36.529.5
Math Index80.386.0
Output speed (tok/s)44.542.7
Blended price ($/1M)$32.81$0.96
Time to first token (s)8.55s1.63s
aime
aime 2580.3%86.0%
artificial analysis coding index36.5029.50
artificial analysis intelligence index42.0032.50
artificial analysis math index80.3086.00
gpqa80.9%78.0%
hle11.9%13.3%
ifbench55.4%43.4%
lcr66.3%54.3%
livecodebench65.4%69.5%
math 500
mmlu pro88.0%82.9%
scicode40.9%38.4%
tau271.4%70.5%
terminalbench hard34.3%25.0%

Benchmark data from Artificial Analysis.