← All comparisons

DeepHermes 3 - Mistral 24B Preview (Non-reasoning) vs Claude 4.1 Opus (Non-reasoning)

Nous Research vs Anthropic — side-by-side benchmark comparison

DeepHermes 3 - Mistral 24B Preview (Non-reasoning)Claude 4.1 Opus (Non-reasoning)
Intelligence Index10.936.0
Coding Index
Math Index
Output speed (tok/s)0.044.7
Blended price ($/1M)$0.00$32.81
Time to first token (s)0.00s1.63s
aime4.7%
aime 25
artificial analysis coding index
artificial analysis intelligence index10.9036.00
artificial analysis math index
gpqa38.2%
hle3.9%
ifbench
lcr
livecodebench19.5%
math 50059.5%
mmlu pro58.0%
scicode22.8%
tau2
terminalbench hard

Benchmark data from Artificial Analysis.