← All comparisons

Claude 4.1 Opus (Non-reasoning) vs Mistral Large 2 (Jul '24)

Anthropic vs Mistral — side-by-side benchmark comparison

Claude 4.1 Opus (Non-reasoning)Mistral Large 2 (Jul '24)
Intelligence Index36.013.0
Coding Index
Math Index0.0
Output speed (tok/s)44.70.0
Blended price ($/1M)$32.81$3.00
Time to first token (s)1.63s0.00s
aime9.3%
aime 250.0%
artificial analysis coding index
artificial analysis intelligence index36.0013.00
artificial analysis math index0.0%
gpqa47.2%
hle3.2%
ifbench31.6%
lcr1.7%
livecodebench26.7%
math 50071.4%
mmlu pro68.3%
scicode27.1%
tau233.0%
terminalbench hard

Benchmark data from Artificial Analysis.