ggml-org/nemotron-3-super-120b-gguf Q4_K GGUF - Free GGUF Download is indexed on GraySoft with repository links, GGUF quant files, and Hugging Face metadata. This page helps you pick a local model for guIDE or other runtimes. See related models in the same shard below.
Model Intelligence Sheet
ggml-org/nemotron-3-super-120b-gguf overview
Nemotron-3-Super-120B GGUF Recommended way to run this model: Then, access http://localhost:8080
Downloads
1,153
Likes
10
Pipeline
—
Library
—
Visibility
Public
Access
Open
Repository Files & Downloads
1 files detected
Direct downloads for all repository files
| File | Type | Quantization | Size | Link |
|---|---|---|---|---|
| Nemotron-3-Super-120B-Q4_K.gguf | GGUF | Q4_K | 65.11 GB | Download |
Model Details Live
Metadata Inspector
Normalized metadata (stored in metadata_json)
{
"metadata": {},
"card_data": {
"base_model": [
"nvidia/Nemotron-3-Super-120B",
"nvidia/NVIDIA-Nemotron-3-Super-120B-A12B-BF16"
],
"frontmatter": {
"base_model": [
"nvidia/Nemotron-3-Super-120B",
"nvidia/NVIDIA-Nemotron-3-Super-120B-A12B-BF16"
]
},
"hero_image_url": "",
"summary": "# Nemotron-3-Super-120B GGUF Recommended way to run this model: ``sh llama-server -hf ggml-org/Nemotron-3-Super-120B-GGUF `` Then, access http://localhost:8080",
"quick_links": [],
"benchmark_table_html": "",
"readme_markdown": "---\nbase_model:\n- nvidia/Nemotron-3-Super-120B\n- nvidia/NVIDIA-Nemotron-3-Super-120B-A12B-BF16\n---\n# Nemotron-3-Super-120B GGUF\n\nRecommended way to run this model:\n\n```sh\nllama-server -hf ggml-org/Nemotron-3-Super-120B-GGUF\n```\n\nThen, access http://localhost:8080",
"related_quantizations": []
},
"tags": [
"gguf",
"base_model:nvidia/NVIDIA-Nemotron-3-Super-120B-A12B-BF16",
"base_model:quantized:nvidia/NVIDIA-Nemotron-3-Super-120B-A12B-BF16",
"endpoints_compatible",
"region:us",
"conversational"
],
"likes": 10,
"downloads": 1153,
"gated": false,
"private": false,
"last_modified": "2026-03-16T19:22:44.000Z",
"created_at": "2026-03-11T18:46:22.000Z",
"pipeline_tag": "",
"library_name": ""
}
Source payload excerpt (from Hugging Face API)
{
"_id": "69b1b87e3fadc91fa2701a27",
"id": "ggml-org/Nemotron-3-Super-120B-GGUF",
"modelId": "ggml-org/Nemotron-3-Super-120B-GGUF",
"sha": "492ce0545407cd8bfc05e543b153c7902bcc7069",
"createdAt": "2026-03-11T18:46:22.000Z",
"lastModified": "2026-03-16T19:22:44.000Z",
"author": "ggml-org",
"downloads": 1153,
"likes": 10,
"gated": false,
"private": false,
"pipeline_tag": "",
"library_name": "",
"siblings_count": 3
}