← All comparisons

Cogito v2.1 (Reasoning) vs Claude 3.5 Sonnet (Oct '24)

Deep Cogito vs Anthropic — side-by-side benchmark comparison

Cogito v2.1 (Reasoning)Claude 3.5 Sonnet (Oct '24)
Intelligence Index15.9
Coding Index24.830.2
Math Index72.7
Output speed (tok/s)80.70.0
Blended price ($/1M)$1.25$6.56
Time to first token (s)0.51s0.00s
aime15.7%
aime 2572.7%
artificial analysis coding index24.8030.20
artificial analysis intelligence index15.90
artificial analysis math index72.70
gpqa76.8%59.9%
hle11.0%3.9%
ifbench46.3%
lcr21.7%
livecodebench68.8%38.1%
math 50077.1%
mmlu pro84.9%77.2%
scicode41.0%36.6%
tau2
terminalbench hard16.7%

Benchmark data from Artificial Analysis.