← All comparisons

Llama 2 Chat 13B vs Claude 4.1 Opus (Reasoning)

Meta vs Anthropic — side-by-side benchmark comparison

Llama 2 Chat 13BClaude 4.1 Opus (Reasoning)
Intelligence Index8.442.0
Coding Index36.5
Math Index80.3
Output speed (tok/s)0.044.5
Blended price ($/1M)$0.00$32.81
Time to first token (s)0.00s8.55s
aime1.7%
aime 2580.3%
artificial analysis coding index36.50
artificial analysis intelligence index8.4042.00
artificial analysis math index80.30
gpqa32.1%80.9%
hle4.7%11.9%
ifbench55.4%
lcr66.3%
livecodebench9.8%65.4%
math 50032.9%
mmlu pro40.6%88.0%
scicode11.8%40.9%
tau271.4%
terminalbench hard34.3%

Benchmark data from Artificial Analysis.