← All comparisons

Claude 3.5 Sonnet (June '24) vs Mistral Large 2 (Nov '24)

Anthropic vs Mistral — side-by-side benchmark comparison

Claude 3.5 Sonnet (June '24)Mistral Large 2 (Nov '24)
Intelligence Index14.215.1
Coding Index26.013.8
Math Index14.0
Output speed (tok/s)0.032.2
Blended price ($/1M)$6.56$3.00
Time to first token (s)0.00s0.60s
aime9.7%11.0%
aime 2514.0%
artificial analysis coding index26.0013.80
artificial analysis intelligence index14.2015.10
artificial analysis math index14.00
gpqa56.0%48.6%
hle3.7%4.0%
ifbench31.2%
lcr5.3%
livecodebench29.3%
math 50069.5%73.6%
mmlu pro75.1%69.7%
scicode31.6%29.2%
tau230.7%
terminalbench hard6.1%

Benchmark data from Artificial Analysis.