GraySoft
Projects Models Compare Cloud benchmarks FAQ Download guIDE →
Model Intelligence Sheet

FedericoFB/Nemotron-3-Nano-Omni-30B-A3B-Reasoning-NVFP4-GGUF overview

A simple conversion made whith llama.cpp convert hf to gguf.py

gguflmstudiolm-studiobase_model:nvidia/Nemotron-3-Nano-Omni-30B-A3B-Reasoning-NVFP4base_model:quantized:nvidia/Nemotron-3-Nano-Omni-30B-A3B-Reasoning-NVFP4license:otherendpoints_compatibleregion:usconversational

Runs locally from ~2.78 GB disk (4 GB VRAM class GPUs with llama.cpp / guIDE).

Downloads
0
Likes
1
Pipeline
—

Repository Files & Downloads

2 GGUF files detected
Direct downloads for local inference
FileTypeQuantizationSizeLink
Nemotron-3-Nano-Omni-30B-A3B-Reasoning-NVFP4-GGUF.ggufGGUFGGUF19.52 GBDownload
mmproj-Nemotron-3-Nano-Omni-30B-A3B-Reasoning-NVFP4-GGUF.ggufGGUFGGUF2.78 GBDownload

Model Details

Model IDFedericoFB/Nemotron-3-Nano-Omni-30B-A3B-Reasoning-NVFP4-GGUF
AuthorFedericoFB
Pipeline—
Licenseother
Base modelnvidia/Nemotron-3-Nano-Omni-30B-A3B-Reasoning-NVFP4
Last modified2026-09-22T09:59:48.000Z

Model README

---

license: other

license_name: nvidia-nemotron-open-model-license

license_link: >-

https://www.nvidia.com/en-us/agreements/enterprise-software/nvidia-nemotron-open-model-license/

base_model: nvidia/Nemotron-3-Nano-Omni-30B-A3B-Reasoning-NVFP4

tags:

- gguf

- lmstudio

- lm-studio

---

A simple conversion made whith llama.cpp convert_hf_to_gguf.py

Run FedericoFB/Nemotron-3-Nano-Omni-30B-A3B-Reasoning-NVFP4-GGUF with guIDE

Download guIDE — the AI-native code editor with local LLM inference and 69 built-in tools.

Download guIDE → · Browse 524k+ models · Compare models

Source: Hugging Face · Compare models