← All comparisons

Claude 4.1 Opus (Reasoning) vs Mistral Small 3.1

Anthropic vs Mistral — side-by-side benchmark comparison

Claude 4.1 Opus (Reasoning)Mistral Small 3.1
Intelligence Index42.014.5
Coding Index36.513.9
Math Index80.33.7
Output speed (tok/s)44.5163.2
Blended price ($/1M)$32.81$0.14
Time to first token (s)8.55s0.49s
aime9.3%
aime 2580.3%3.7%
artificial analysis coding index36.5013.90
artificial analysis intelligence index42.0014.50
artificial analysis math index80.303.70
gpqa80.9%45.4%
hle11.9%4.8%
ifbench55.4%29.9%
lcr66.3%19.7%
livecodebench65.4%21.2%
math 50070.7%
mmlu pro88.0%65.9%
scicode40.9%26.5%
tau271.4%25.1%
terminalbench hard34.3%7.6%

Benchmark data from Artificial Analysis.