← All comparisons

Claude 3.5 Sonnet (June '24) vs GLM-4.7 (Reasoning)

Anthropic vs Z AI — side-by-side benchmark comparison

Claude 3.5 Sonnet (June '24)GLM-4.7 (Reasoning)
Intelligence Index14.242.1
Coding Index26.036.3
Math Index95.0
Output speed (tok/s)0.079.6
Blended price ($/1M)$6.56$1.00
Time to first token (s)0.00s0.73s
aime9.7%
aime 2595.0%
artificial analysis coding index26.0036.30
artificial analysis intelligence index14.2042.10
artificial analysis math index95.00
gpqa56.0%85.9%
hle3.7%25.1%
ifbench67.9%
lcr64.0%
livecodebench89.4%
math 50069.5%
mmlu pro75.1%85.6%
scicode31.6%45.1%
tau295.9%
terminalbench hard31.8%

Benchmark data from Artificial Analysis.