← All comparisons

GPT-4.1 mini vs Claude 4.1 Opus (Reasoning)

OpenAI vs Anthropic — side-by-side benchmark comparison

GPT-4.1 miniClaude 4.1 Opus (Reasoning)
Intelligence Index22.942.0
Coding Index18.536.5
Math Index46.380.3
Output speed (tok/s)95.144.5
Blended price ($/1M)$0.70$32.81
Time to first token (s)0.59s8.55s
aime43.0%
aime 2546.3%80.3%
artificial analysis coding index18.5036.50
artificial analysis intelligence index22.9042.00
artificial analysis math index46.3080.30
gpqa66.4%80.9%
hle4.6%11.9%
ifbench38.3%55.4%
lcr42.3%66.3%
livecodebench48.3%65.4%
math 50092.5%
mmlu pro78.1%88.0%
scicode40.4%40.9%
tau252.9%71.4%
terminalbench hard7.6%34.3%

Benchmark data from Artificial Analysis.