GraySoft
Projects Models Compare Cloud benchmarks FAQ Download guIDE →

NVIDIA cloud models

19 models tracked via Artificial Analysis. Compare cloud performance, then find local GGUF versions in the GraySoft model catalog.

ModelIntelligenceSpeed (tok/s)
Nemotron 3 Ultra 550B A55B (Reasoning)23.4175.767
Nemotron 3.5 Lightning13.6276.957
Nemotron 3 Super 120B A12B (Reasoning)13.6105.583
Nemotron Cascade 2 30B A3B11.70
Nemotron 3 Nano Omni 30B A3B Reasoning10.30
Llama Nemotron Super 49B v1.5 (Reasoning)958.706
NVIDIA Nemotron 3 Nano 30B A3B (Reasoning)8.9231.825
Llama 3.3 Nemotron Super 49B v1 (Reasoning)8.90
Llama 3.1 Nemotron Ultra 253B v1 (Reasoning)7.50
NVIDIA Nemotron Nano 12B v2 VL (Reasoning)7.568.502
NVIDIA Nemotron Nano 9B V2 (Reasoning)7.4100.118
Llama Nemotron Super 49B v1.5 (Non-reasoning)7.451.119
NVIDIA Nemotron 3 Nano 4B7.40
Llama 3.3 Nemotron Super 49B v1 (Non-reasoning)7.30
Llama 3.1 Nemotron Nano 4B v1.1 (Reasoning)7.30
Llama 3.1 Nemotron Instruct 70B6.963.28
NVIDIA Nemotron Nano 9B V2 (Non-reasoning)6.8153.738
NVIDIA Nemotron 3 Nano 30B A3B (Non-reasoning)6.8178.472
NVIDIA Nemotron Nano 12B v2 VL (Non-reasoning)5.872.209

Run models locally with guIDE

Download guIDE — the AI-native code editor with local LLM inference and 69 built-in tools.

Download guIDE → · Browse 524k+ models · Compare models