← All comparisons

Gemini 2.5 Flash Preview (Sep '25) (Non-reasoning) vs Claude 4.5 Sonnet (Reasoning)

Google vs Anthropic — side-by-side benchmark comparison

Gemini 2.5 Flash Preview (Sep '25) (Non-reasoning)Claude 4.5 Sonnet (Reasoning)
Intelligence Index25.743.0
Coding Index22.138.6
Math Index56.788.0
Output speed (tok/s)0.055.0
Blended price ($/1M)$0.00$6.56
Time to first token (s)0.00s7.02s
aime
aime 2556.7%88.0%
artificial analysis coding index22.1038.60
artificial analysis intelligence index25.7043.00
artificial analysis math index56.7088.00
gpqa76.6%83.4%
hle7.8%17.3%
ifbench43.5%57.3%
lcr56.7%65.7%
livecodebench62.5%71.4%
math 500
mmlu pro83.6%87.5%
scicode37.5%44.7%
tau228.4%78.1%
terminalbench hard14.4%35.6%

Benchmark data from Artificial Analysis.