← All comparisons

Claude 3.5 Sonnet (Oct '24) vs Grok 3 mini Reasoning (high)

Anthropic vs xAI — side-by-side benchmark comparison

Claude 3.5 Sonnet (Oct '24)Grok 3 mini Reasoning (high)
Intelligence Index15.932.1
Coding Index30.225.2
Math Index84.7
Output speed (tok/s)0.056.8
Blended price ($/1M)$6.56$0.35
Time to first token (s)0.00s0.42s
aime15.7%93.3%
aime 2584.7%
artificial analysis coding index30.2025.20
artificial analysis intelligence index15.9032.10
artificial analysis math index84.70
gpqa59.9%79.1%
hle3.9%11.1%
ifbench45.9%
lcr50.3%
livecodebench38.1%69.6%
math 50077.1%99.2%
mmlu pro77.2%82.8%
scicode36.6%40.6%
tau290.4%
terminalbench hard17.4%

Benchmark data from Artificial Analysis.