← All comparisons

Claude 4.1 Opus (Non-reasoning) vs Grok 3 mini Reasoning (high)

Anthropic vs xAI — side-by-side benchmark comparison

Claude 4.1 Opus (Non-reasoning)Grok 3 mini Reasoning (high)
Intelligence Index36.032.1
Coding Index25.2
Math Index84.7
Output speed (tok/s)44.756.8
Blended price ($/1M)$32.81$0.35
Time to first token (s)1.63s0.42s
aime93.3%
aime 2584.7%
artificial analysis coding index25.20
artificial analysis intelligence index36.0032.10
artificial analysis math index84.70
gpqa79.1%
hle11.1%
ifbench45.9%
lcr50.3%
livecodebench69.6%
math 50099.2%
mmlu pro82.8%
scicode40.6%
tau290.4%
terminalbench hard17.4%

Benchmark data from Artificial Analysis.