← All comparisons

Claude 3.5 Sonnet (June '24) vs GLM-4.7 (Non-reasoning)

Anthropic vs Z AI — side-by-side benchmark comparison

Claude 3.5 Sonnet (June '24)GLM-4.7 (Non-reasoning)
Intelligence Index14.234.2
Coding Index26.032.0
Math Index48.0
Output speed (tok/s)0.084.8
Blended price ($/1M)$6.56$1.00
Time to first token (s)0.00s0.75s
aime9.7%
aime 2548.0%
artificial analysis coding index26.0032.00
artificial analysis intelligence index14.2034.20
artificial analysis math index48.00
gpqa56.0%66.4%
hle3.7%6.1%
ifbench54.6%
lcr36.3%
livecodebench56.2%
math 50069.5%
mmlu pro75.1%79.4%
scicode31.6%35.4%
tau294.2%
terminalbench hard30.3%

Benchmark data from Artificial Analysis.