ggml-org/llama-4-scout-17b-16e-instruct-gguf F16 GGUF - Free GGUF Download is indexed on GraySoft with repository links, GGUF quant files, and Hugging Face metadata. This page helps you pick a local model for guIDE or other runtimes. See related models in the same shard below.
Model Intelligence Sheet
ggml-org/llama-4-scout-17b-16e-instruct-gguf overview
Related to this PR: https://github.com/ggml-org/llama.cpp/pull/13282 Quantizations for text model are taken from unsloth, all credits to them!
Downloads
518
Likes
7
Pipeline
—
Library
—
Visibility
Public
Access
Open
Repository Files & Downloads
4 files detected
Direct downloads for all repository files
| File | Type | Quantization | Size | Link |
|---|---|---|---|---|
| Llama-4-Scout-17B-16E-Instruct-Q4_K_M-00001-of-00002.gguf | GGUF | Q4_K_M | 46.42 GB | Download |
| Llama-4-Scout-17B-16E-Instruct-Q4_K_M-00002-of-00002.gguf | GGUF | Q4_K_M | 14.45 GB | Download |
| Llama-4-Scout-17B-16E-Instruct-UD-IQ1_S.gguf | GGUF | IQ1_S | 30.24 GB | Download |
| mmproj-Llama-4-Scout-17B-16E-Instruct-f16.gguf | GGUF | F16 | 1.63 GB | Download |
Model Details Live
Metadata Inspector
Normalized metadata (stored in metadata_json)
{
"metadata": {},
"card_data": {
"base_model": [
"meta-llama/Llama-4-Scout-17B-16E-Instruct"
],
"frontmatter": {
"base_model": [
"meta-llama/Llama-4-Scout-17B-16E-Instruct"
]
},
"hero_image_url": "",
"summary": "Related to this PR: https://github.com/ggml-org/llama.cpp/pull/13282 Quantizations for **text** model are taken from unsloth, all credits to them!",
"quick_links": [],
"benchmark_table_html": "",
"readme_markdown": "---\nbase_model:\n- meta-llama/Llama-4-Scout-17B-16E-Instruct\n---\n\n# Llama 4 Scout with vision support\n\nRelated to this PR: https://github.com/ggml-org/llama.cpp/pull/13282\n\nQuantizations for **text** model are taken from [unsloth](https://huggingface.co/unsloth/Llama-4-Scout-17B-16E-Instruct-GGUF), all credits to them! \n",
"related_quantizations": []
},
"tags": [
"gguf",
"base_model:meta-llama/Llama-4-Scout-17B-16E-Instruct",
"base_model:quantized:meta-llama/Llama-4-Scout-17B-16E-Instruct",
"endpoints_compatible",
"region:us",
"imatrix",
"conversational"
],
"likes": 7,
"downloads": 518,
"gated": false,
"private": false,
"last_modified": "2025-05-18T14:34:30.000Z",
"created_at": "2025-05-18T13:58:19.000Z",
"pipeline_tag": "",
"library_name": ""
}
Source payload excerpt (from Hugging Face API)
{
"_id": "6829e77b73127d39651a5ba1",
"id": "ggml-org/Llama-4-Scout-17B-16E-Instruct-GGUF",
"modelId": "ggml-org/Llama-4-Scout-17B-16E-Instruct-GGUF",
"sha": "42675345da11ade9203a5187595da7b74d4ff2ac",
"createdAt": "2025-05-18T13:58:19.000Z",
"lastModified": "2025-05-18T14:34:30.000Z",
"author": "ggml-org",
"downloads": 518,
"likes": 7,
"gated": false,
"private": false,
"pipeline_tag": "",
"library_name": "",
"siblings_count": 6
}