← All comparisons

Claude 4.5 Haiku (Reasoning) vs GPT-5.1 Codex mini (high)

Anthropic vs OpenAI — side-by-side benchmark comparison

Claude 4.5 Haiku (Reasoning)GPT-5.1 Codex mini (high)
Intelligence Index37.138.6
Coding Index32.636.4
Math Index83.791.7
Output speed (tok/s)142.2218.5
Blended price ($/1M)$2.19$0.69
Time to first token (s)10.48s3.35s
aime
aime 2583.7%91.7%
artificial analysis coding index32.6036.40
artificial analysis intelligence index37.1038.60
artificial analysis math index83.7091.70
gpqa67.2%81.3%
hle9.7%16.9%
ifbench54.3%67.9%
lcr70.3%62.7%
livecodebench61.5%83.6%
math 500
mmlu pro76.0%82.0%
scicode43.3%42.6%
tau254.7%62.9%
terminalbench hard27.3%33.3%

Benchmark data from Artificial Analysis.