← All comparisons

Claude 3.5 Sonnet (June '24) vs GLM-4.6 (Non-reasoning)

Anthropic vs Z AI — side-by-side benchmark comparison

Claude 3.5 Sonnet (June '24)GLM-4.6 (Non-reasoning)
Intelligence Index14.230.2
Coding Index26.030.2
Math Index44.3
Output speed (tok/s)0.053.2
Blended price ($/1M)$6.56$1.00
Time to first token (s)0.00s1.84s
aime9.7%
aime 2544.3%
artificial analysis coding index26.0030.20
artificial analysis intelligence index14.2030.20
artificial analysis math index44.30
gpqa56.0%63.2%
hle3.7%5.2%
ifbench36.7%
lcr26.3%
livecodebench56.1%
math 50069.5%
mmlu pro75.1%78.4%
scicode31.6%33.1%
tau276.9%
terminalbench hard28.8%

Benchmark data from Artificial Analysis.