← All comparisons

Claude 3.5 Sonnet (Oct '24) vs Mistral Large 2 (Nov '24)

Anthropic vs Mistral — side-by-side benchmark comparison

Claude 3.5 Sonnet (Oct '24)Mistral Large 2 (Nov '24)
Intelligence Index15.915.1
Coding Index30.213.8
Math Index14.0
Output speed (tok/s)0.032.2
Blended price ($/1M)$6.56$3.00
Time to first token (s)0.00s0.60s
aime15.7%11.0%
aime 2514.0%
artificial analysis coding index30.2013.80
artificial analysis intelligence index15.9015.10
artificial analysis math index14.00
gpqa59.9%48.6%
hle3.9%4.0%
ifbench31.2%
lcr5.3%
livecodebench38.1%29.3%
math 50077.1%73.6%
mmlu pro77.2%69.7%
scicode36.6%29.2%
tau230.7%
terminalbench hard6.1%

Benchmark data from Artificial Analysis.