← All comparisons

Gemini 2.0 Flash Thinking Experimental (Jan '25) vs Claude 4.5 Sonnet (Reasoning)

Google vs Anthropic — side-by-side benchmark comparison

Gemini 2.0 Flash Thinking Experimental (Jan '25)Claude 4.5 Sonnet (Reasoning)
Intelligence Index19.643.0
Coding Index24.138.6
Math Index88.0
Output speed (tok/s)0.055.0
Blended price ($/1M)$0.00$6.56
Time to first token (s)0.00s7.02s
aime50.0%
aime 2588.0%
artificial analysis coding index24.1038.60
artificial analysis intelligence index19.6043.00
artificial analysis math index88.00
gpqa70.1%83.4%
hle7.1%17.3%
ifbench57.3%
lcr65.7%
livecodebench32.1%71.4%
math 50094.4%
mmlu pro79.8%87.5%
scicode32.9%44.7%
tau278.1%
terminalbench hard35.6%

Benchmark data from Artificial Analysis.