spiritfather/Gemma4-Gutenberg-26B-A4B-i1-GGUF overview
About weighted/imatrix quants of https://huggingface.co/nbeerbower/Gemma4 Gutenberg 26B A4B Note: Gemma4 Gutenberg 26B A4B https://huggingface.co/nbeerbower/Ge…
Runs locally from ~7.72 GB disk (8 GB VRAM class GPUs with llama.cpp / guIDE).
Repository Files & Downloads
| File | Type | Quantization | Size | Link |
|---|---|---|---|---|
| Gemma4-Gutenberg-26B-A4B.i1-IQ1_M.gguf | GGUF | IQ1_M | 8.07 GB | Download |
| Gemma4-Gutenberg-26B-A4B.i1-IQ1_S.gguf | GGUF | IQ1_S | 7.72 GB | Download |
| Gemma4-Gutenberg-26B-A4B.i1-IQ2_M.gguf | GGUF | IQ2_M | 9.67 GB | Download |
| Gemma4-Gutenberg-26B-A4B.i1-IQ2_S.gguf | GGUF | IQ2_S | 9.20 GB | Download |
| Gemma4-Gutenberg-26B-A4B.i1-IQ2_XS.gguf | GGUF | IQ2_XS | 9.14 GB | Download |
| Gemma4-Gutenberg-26B-A4B.i1-IQ2_XXS.gguf | GGUF | IQ2_XXS | 8.66 GB | Download |
| Gemma4-Gutenberg-26B-A4B.i1-IQ3_M.gguf | GGUF | IQ3_M | 11.54 GB | Download |
| Gemma4-Gutenberg-26B-A4B.i1-IQ3_S.gguf | GGUF | IQ3_S | 11.38 GB | Download |
| Gemma4-Gutenberg-26B-A4B.i1-IQ3_XS.gguf | GGUF | IQ3_XS | 10.84 GB | Download |
| Gemma4-Gutenberg-26B-A4B.i1-IQ3_XXS.gguf | GGUF | IQ3_XXS | 10.55 GB | Download |
| Gemma4-Gutenberg-26B-A4B.i1-IQ4_XS.gguf | GGUF | IQ4_XS | 12.96 GB | Download |
| Gemma4-Gutenberg-26B-A4B.i1-Q2_K.gguf | GGUF | Q2_K | 9.86 GB | Download |
| Gemma4-Gutenberg-26B-A4B.i1-Q2_K_S.gguf | GGUF | Q2_K_S | 9.89 GB | Download |
| Gemma4-Gutenberg-26B-A4B.i1-Q3_K_L.gguf | GGUF | Q3_K_L | 12.88 GB | Download |
| Gemma4-Gutenberg-26B-A4B.i1-Q3_K_M.gguf | GGUF | Q3_K_M | 12.37 GB | Download |
| Gemma4-Gutenberg-26B-A4B.i1-Q3_K_S.gguf | GGUF | Q3_K_S | 11.38 GB | Download |
| Gemma4-Gutenberg-26B-A4B.i1-Q4_0.gguf | GGUF | Q4_0 | 13.49 GB | Download |
| Gemma4-Gutenberg-26B-A4B.i1-Q4_1.gguf | GGUF | Q4_1 | 14.87 GB | Download |
| Gemma4-Gutenberg-26B-A4B.i1-Q4_K_M.gguf | GGUF | Q4_K_M | 15.64 GB | Download |
| Gemma4-Gutenberg-26B-A4B.i1-Q4_K_S.gguf | GGUF | Q4_K_S | 14.40 GB | Download |
| Gemma4-Gutenberg-26B-A4B.i1-Q5_K_M.gguf | GGUF | Q5_K_M | 17.82 GB | Download |
| Gemma4-Gutenberg-26B-A4B.i1-Q5_K_S.gguf | GGUF | Q5_K_S | 16.75 GB | Download |
| Gemma4-Gutenberg-26B-A4B.i1-Q6_K.gguf | GGUF | Q6_K | 21.08 GB | Download |
Model Details
| Model ID | spiritfather/Gemma4-Gutenberg-26B-A4B-i1-GGUF |
|---|---|
| Author | spiritfather |
| Pipeline | text-generation |
| License | gemma |
| Base model | nbeerbower/Gemma4-Gutenberg-26B-A4B |
| Last modified | 2026-07-07T15:40:55.000Z |
Model README
---
base_model: nbeerbower/Gemma4-Gutenberg-26B-A4B
language:
- en
library_name: transformers
license: gemma
quantized_by: spiritfather
tags:
- text-generation
- creative-writing
- gguf
- imatrix
---
About
weighted/imatrix quants of https://huggingface.co/nbeerbower/Gemma4-Gutenberg-26B-A4B
> Note: Gemma4-Gutenberg-26B-A4B is published as a LoRA adapter, not a full model. These GGUFs are that adapter merged onto its base google/gemma-4-26B-A4B-it (the text LM extracted from the multimodal base), then quantized.
> imatrix (importance matrix) computed over bartowski calibration_datav3. Weighted quants spend precision where it matters most — the low bit-rates (IQ1–IQ4) are meaningfully better than same-size static quants; at Q5/Q6 the difference is negligible.
> Static (non-imatrix) quants of this model are at spiritfather/Gemma4-Gutenberg-26B-A4B-GGUF.
<!-- provided-files -->
Usage
If you are unsure how to use GGUF files, refer to one of TheBloke's READMEs for more details, including on how to concatenate multi-part files.
Provided Quants
(sorted by size, not necessarily quality. IQ-quants are often preferable over similar sized non-IQ quants)
| Link | Type | Size/GB | Notes |
|:-----|:-----|--------:|:------|
| GGUF | IQ1_S | 8.3 | for the desperate |
| GGUF | IQ1_M | 8.7 | mostly desperate |
| GGUF | IQ2_XXS | 9.3 | lower quality |
| GGUF | IQ2_XS | 9.8 | |
| GGUF | IQ2_S | 9.9 | |
| GGUF | IQ2_M | 10.4 | |
| GGUF | Q2_K_S | 10.6 | |
| GGUF | Q2_K | 10.6 | |
| GGUF | IQ3_XXS | 11.3 | lower quality |
| GGUF | IQ3_XS | 11.6 | |
| GGUF | Q3_K_S | 12.2 | |
| GGUF | IQ3_S | 12.2 | |
| GGUF | IQ3_M | 12.4 | |
| GGUF | Q3_K_M | 13.3 | IQ3_S probably better |
| GGUF | Q3_K_L | 13.8 | |
| GGUF | IQ4_XS | 13.9 | |
| GGUF | Q4_0 | 14.5 | |
| GGUF | Q4_K_S | 15.5 | optimal size/speed/quality |
| GGUF | Q4_K_M | 16.8 | fast, recommended |
| GGUF | Q4_1 | 16.0 | |
| GGUF | Q5_K_S | 18.0 | |
| GGUF | Q5_K_M | 19.1 | |
| GGUF | Q6_K | 22.6 | practically like static Q6_K |
Run spiritfather/Gemma4-Gutenberg-26B-A4B-i1-GGUF with guIDE
Download guIDE — the AI-native code editor with local LLM inference and 69 built-in tools.
Source: Hugging Face · Compare models