← All comparisons

Claude 4.1 Opus (Reasoning) vs Mistral Small 3.2

Anthropic vs Mistral — side-by-side benchmark comparison

Claude 4.1 Opus (Reasoning)Mistral Small 3.2
Intelligence Index42.015.1
Coding Index36.513.3
Math Index80.327.0
Output speed (tok/s)44.5133.0
Blended price ($/1M)$32.81$0.13
Time to first token (s)8.55s0.36s
aime32.3%
aime 2580.3%27.0%
artificial analysis coding index36.5013.30
artificial analysis intelligence index42.0015.10
artificial analysis math index80.3027.00
gpqa80.9%50.5%
hle11.9%4.3%
ifbench55.4%33.5%
lcr66.3%17.3%
livecodebench65.4%27.5%
math 50088.3%
mmlu pro88.0%68.1%
scicode40.9%26.4%
tau271.4%29.5%
terminalbench hard34.3%6.8%

Benchmark data from Artificial Analysis.