← All comparisons

Claude 4.1 Opus (Reasoning) vs Grok 4

Anthropic vs xAI — side-by-side benchmark comparison

Claude 4.1 Opus (Reasoning)Grok 4
Intelligence Index42.041.5
Coding Index36.540.5
Math Index80.392.7
Output speed (tok/s)44.50.0
Blended price ($/1M)$32.81$11.00
Time to first token (s)8.55s0.00s
aime94.3%
aime 2580.3%92.7%
artificial analysis coding index36.5040.50
artificial analysis intelligence index42.0041.50
artificial analysis math index80.3092.70
gpqa80.9%87.7%
hle11.9%23.9%
ifbench55.4%53.7%
lcr66.3%68.0%
livecodebench65.4%81.9%
math 50099.0%
mmlu pro88.0%86.6%
scicode40.9%45.7%
tau271.4%74.9%
terminalbench hard34.3%37.9%

Benchmark data from Artificial Analysis.