← All comparisons

Gemma 4 31B (Reasoning) vs Claude 4.5 Haiku (Non-reasoning)

Google vs Anthropic — side-by-side benchmark comparison

Gemma 4 31B (Reasoning)Claude 4.5 Haiku (Non-reasoning)
Intelligence Index39.231.0
Coding Index38.729.6
Math Index39.0
Output speed (tok/s)35.3105.8
Blended price ($/1M)$0.00$2.19
Time to first token (s)1.00s0.65s
aime
aime 2539.0%
artificial analysis coding index38.7029.60
artificial analysis intelligence index39.2031.00
artificial analysis math index39.00
gpqa85.7%64.6%
hle22.7%4.3%
ifbench75.6%42.0%
lcr62.0%43.7%
livecodebench51.1%
math 500
mmlu pro80.0%
scicode43.4%34.4%
tau259.9%32.5%
terminalbench hard36.4%27.3%

Benchmark data from Artificial Analysis.