← All comparisons

Claude 4.1 Opus (Reasoning) vs Mistral Small 3

Anthropic vs Mistral — side-by-side benchmark comparison

Claude 4.1 Opus (Reasoning)Mistral Small 3
Intelligence Index42.012.7
Coding Index36.5
Math Index80.34.3
Output speed (tok/s)44.5150.2
Blended price ($/1M)$32.81$0.10
Time to first token (s)8.55s0.48s
aime8.0%
aime 2580.3%4.3%
artificial analysis coding index36.50
artificial analysis intelligence index42.0012.70
artificial analysis math index80.304.30
gpqa80.9%46.2%
hle11.9%4.1%
ifbench55.4%26.4%
lcr66.3%0.0%
livecodebench65.4%25.2%
math 50071.5%
mmlu pro88.0%65.2%
scicode40.9%23.6%
tau271.4%19.6%
terminalbench hard34.3%

Benchmark data from Artificial Analysis.