GraySoft
Projects Models Compare Cloud benchmarks FAQ Download guIDE →
Model Intelligence Sheet

ggml-org/NVIDIA-Nemotron-3.5-Lightning-30B-A3B-GGUF overview

NVIDIA Nemotron 3.5 Lightning 30B A3B Run with https://llama.app bash llama serve hf ggml org/NVIDIA Nemotron 3.5 Lightning 30B A3B GGUF Source models https://…

ggufquantizedtext-generationbase_model:nvidia/NVIDIA-Nemotron-3.5-Lightning-30B-A3B-BF16base_model:quantized:nvidia/NVIDIA-Nemotron-3.5-Lightning-30B-A3B-BF16license:otherendpoints_compatibleregion:usconversational

Runs locally from ~1.08 GB disk (4 GB VRAM class GPUs with llama.cpp / guIDE).

Downloads
117,506
Likes
25
Pipeline
text-generation
Author

Repository Files & Downloads

6 GGUF files detected
Direct downloads for local inference
FileTypeQuantizationSizeLink
NVIDIA-Nemotron-3.5-Lightning-30B-A3B-BF16.ggufGGUFBF1658.84 GBDownload
NVIDIA-Nemotron-3.5-Lightning-30B-A3B-Q4_0.ggufGGUFQ4_017.60 GBDownload
NVIDIA-Nemotron-3.5-Lightning-30B-A3B-Q8_0.ggufGGUFQ8_031.28 GBDownload
mtp-NVIDIA-Nemotron-3.5-Lightning-30B-A3B-BF16.ggufGGUFBF163.81 GBDownload
mtp-NVIDIA-Nemotron-3.5-Lightning-30B-A3B-Q4_0.ggufGGUFQ4_01.08 GBDownload
mtp-NVIDIA-Nemotron-3.5-Lightning-30B-A3B-Q8_0.ggufGGUFQ8_02.03 GBDownload

Model Details

Model IDggml-org/NVIDIA-Nemotron-3.5-Lightning-30B-A3B-GGUF
Authorggml-org
Pipelinetext-generation
Licenseother
Base modelnvidia/NVIDIA-Nemotron-3.5-Lightning-30B-A3B-BF16
Last modified2026-08-23T02:18:10.000Z

Model README

---

license: other

pipeline_tag: text-generation

tags:

  • gguf
  • quantized

base_model:

  • nvidia/NVIDIA-Nemotron-3.5-Lightning-30B-A3B-BF16

---

NVIDIA-Nemotron-3.5-Lightning-30B-A3B

Run with https://llama.app

llama serve -hf ggml-org/NVIDIA-Nemotron-3.5-Lightning-30B-A3B-GGUF

Source models

  • https://huggingface.co/nvidia/NVIDIA-Nemotron-3.5-Lightning-30B-A3B-BF16

<!--

  • https://huggingface.co/nvidia/NVIDIA-Nemotron-3.5-Lightning-30B-A3B-NVFP4
  • https://huggingface.co/nvidia/NVIDIA-Nemotron-3.5-Lightning-30B-A3B-NVFP4-DFlash
  • https://huggingface.co/nvidia/NVIDIA-Nemotron-3.5-Lightning-30B-A3B-NVFP4-DSpark

-->

TODOs

  • add info

> [!IMPORTANT]

> This model is automatically converted using https://github.com/ggml-org/convert

Run ggml-org/NVIDIA-Nemotron-3.5-Lightning-30B-A3B-GGUF with guIDE

Download guIDE — the AI-native code editor with local LLM inference and 69 built-in tools.

Download guIDE → · Browse 524k+ models · Compare models

Source: Hugging Face · Compare models