Model Intelligence Sheet
Xlr8boi/Llama-3.2-1B-GGUF overview
This is the gguf repo for the model Llama 3.2 1B, having all the quantized models.
Runs locally from ~554.0 MB disk (4 GB VRAM class GPUs with llama.cpp / guIDE).
Repository Files & Downloads
19 GGUF files detected
Direct downloads for local inference
| File | Type | Quantization | Size | Link |
|---|---|---|---|---|
| Llama-3.2-1B-Q4_K_M.gguf | GGUF | Q4_K_M | 770.3 MB | Download |
| Llama-3.2-1B-f16.gguf | GGUF | F16 | 2.31 GB | Download |
| llama-IQ3_M.gguf | GGUF | IQ3_M | 626.8 MB | Download |
| llama-IQ3_S.gguf | GGUF | IQ3_S | 614.1 MB | Download |
| llama-IQ4_NL.gguf | GGUF | IQ4_NL | 741.2 MB | Download |
| llama-IQ4_XS.gguf | GGUF | IQ4_XS | 713.7 MB | Download |
| llama-Q2_K.gguf | GGUF | Q2_K | 554.0 MB | Download |
| llama-Q3_K_L.gguf | GGUF | Q3_K_L | 698.6 MB | Download |
| llama-Q3_K_M.gguf | GGUF | Q3_K_M | 658.8 MB | Download |
| llama-Q3_K_S.gguf | GGUF | Q3_K_S | 612.0 MB | Download |
| llama-Q4_0.gguf | GGUF | Q4_0 | 735.2 MB | Download |
| llama-Q4_1.gguf | GGUF | Q4_1 | 793.2 MB | Download |
| llama-Q4_K_S.gguf | GGUF | Q4_K_S | 739.7 MB | Download |
| llama-Q5_0.gguf | GGUF | Q5_0 | 851.2 MB | Download |
| llama-Q5_1.gguf | GGUF | Q5_1 | 909.2 MB | Download |
| llama-Q5_K_M.gguf | GGUF | Q5_K_M | 869.3 MB | Download |
| llama-Q5_K_S.gguf | GGUF | Q5_K_S | 851.2 MB | Download |
| llama-Q6_K.gguf | GGUF | Q6_K | 974.5 MB | Download |
| llama-Q8_0.gguf | GGUF | Q8_0 | 1.23 GB | Download |
Model Details
Model README
This is the gguf repo for the model Llama-3.2-1B, having all the quantized models.
Run Xlr8boi/Llama-3.2-1B-GGUF with guIDE
Download guIDE — the AI-native code editor with local LLM inference and 69 built-in tools.
Source: Hugging Face · Compare models