GraySoft
Projects Models Compare Cloud benchmarks FAQ Download guIDE →
Model Intelligence Sheet

spiritfather/Gemma4-Gutenberg-26B-A4B-i1-GGUF overview

About weighted/imatrix quants of https://huggingface.co/nbeerbower/Gemma4 Gutenberg 26B A4B Note: Gemma4 Gutenberg 26B A4B https://huggingface.co/nbeerbower/Ge…

transformersgguftext-generationcreative-writingimatrixenbase_model:nbeerbower/Gemma4-Gutenberg-26B-A4Bbase_model:quantized:nbeerbower/Gemma4-Gutenberg-26B-A4Blicense:gemmaendpoints_compatibleregion:usconversational

Runs locally from ~7.72 GB disk (8 GB VRAM class GPUs with llama.cpp / guIDE).

Downloads
0
Likes
0
Pipeline
text-generation

Repository Files & Downloads

23 GGUF files detected
Direct downloads for local inference
FileTypeQuantizationSizeLink
Gemma4-Gutenberg-26B-A4B.i1-IQ1_M.ggufGGUFIQ1_M8.07 GBDownload
Gemma4-Gutenberg-26B-A4B.i1-IQ1_S.ggufGGUFIQ1_S7.72 GBDownload
Gemma4-Gutenberg-26B-A4B.i1-IQ2_M.ggufGGUFIQ2_M9.67 GBDownload
Gemma4-Gutenberg-26B-A4B.i1-IQ2_S.ggufGGUFIQ2_S9.20 GBDownload
Gemma4-Gutenberg-26B-A4B.i1-IQ2_XS.ggufGGUFIQ2_XS9.14 GBDownload
Gemma4-Gutenberg-26B-A4B.i1-IQ2_XXS.ggufGGUFIQ2_XXS8.66 GBDownload
Gemma4-Gutenberg-26B-A4B.i1-IQ3_M.ggufGGUFIQ3_M11.54 GBDownload
Gemma4-Gutenberg-26B-A4B.i1-IQ3_S.ggufGGUFIQ3_S11.38 GBDownload
Gemma4-Gutenberg-26B-A4B.i1-IQ3_XS.ggufGGUFIQ3_XS10.84 GBDownload
Gemma4-Gutenberg-26B-A4B.i1-IQ3_XXS.ggufGGUFIQ3_XXS10.55 GBDownload
Gemma4-Gutenberg-26B-A4B.i1-IQ4_XS.ggufGGUFIQ4_XS12.96 GBDownload
Gemma4-Gutenberg-26B-A4B.i1-Q2_K.ggufGGUFQ2_K9.86 GBDownload
Gemma4-Gutenberg-26B-A4B.i1-Q2_K_S.ggufGGUFQ2_K_S9.89 GBDownload
Gemma4-Gutenberg-26B-A4B.i1-Q3_K_L.ggufGGUFQ3_K_L12.88 GBDownload
Gemma4-Gutenberg-26B-A4B.i1-Q3_K_M.ggufGGUFQ3_K_M12.37 GBDownload
Gemma4-Gutenberg-26B-A4B.i1-Q3_K_S.ggufGGUFQ3_K_S11.38 GBDownload
Gemma4-Gutenberg-26B-A4B.i1-Q4_0.ggufGGUFQ4_013.49 GBDownload
Gemma4-Gutenberg-26B-A4B.i1-Q4_1.ggufGGUFQ4_114.87 GBDownload
Gemma4-Gutenberg-26B-A4B.i1-Q4_K_M.ggufGGUFQ4_K_M15.64 GBDownload
Gemma4-Gutenberg-26B-A4B.i1-Q4_K_S.ggufGGUFQ4_K_S14.40 GBDownload
Gemma4-Gutenberg-26B-A4B.i1-Q5_K_M.ggufGGUFQ5_K_M17.82 GBDownload
Gemma4-Gutenberg-26B-A4B.i1-Q5_K_S.ggufGGUFQ5_K_S16.75 GBDownload
Gemma4-Gutenberg-26B-A4B.i1-Q6_K.ggufGGUFQ6_K21.08 GBDownload

Model Details

Model IDspiritfather/Gemma4-Gutenberg-26B-A4B-i1-GGUF
Authorspiritfather
Pipelinetext-generation
Licensegemma
Base modelnbeerbower/Gemma4-Gutenberg-26B-A4B
Last modified2026-07-07T15:40:55.000Z

Model README

---

base_model: nbeerbower/Gemma4-Gutenberg-26B-A4B

language:

  • en

library_name: transformers

license: gemma

quantized_by: spiritfather

tags:

  • text-generation
  • creative-writing
  • gguf
  • imatrix

---

About

weighted/imatrix quants of https://huggingface.co/nbeerbower/Gemma4-Gutenberg-26B-A4B

> Note: Gemma4-Gutenberg-26B-A4B is published as a LoRA adapter, not a full model. These GGUFs are that adapter merged onto its base google/gemma-4-26B-A4B-it (the text LM extracted from the multimodal base), then quantized.

> imatrix (importance matrix) computed over bartowski calibration_datav3. Weighted quants spend precision where it matters most — the low bit-rates (IQ1–IQ4) are meaningfully better than same-size static quants; at Q5/Q6 the difference is negligible.

> Static (non-imatrix) quants of this model are at spiritfather/Gemma4-Gutenberg-26B-A4B-GGUF.

<!-- provided-files -->

Usage

If you are unsure how to use GGUF files, refer to one of TheBloke's READMEs for more details, including on how to concatenate multi-part files.

Provided Quants

(sorted by size, not necessarily quality. IQ-quants are often preferable over similar sized non-IQ quants)

| Link | Type | Size/GB | Notes |

|:-----|:-----|--------:|:------|

| GGUF | IQ1_S | 8.3 | for the desperate |

| GGUF | IQ1_M | 8.7 | mostly desperate |

| GGUF | IQ2_XXS | 9.3 | lower quality |

| GGUF | IQ2_XS | 9.8 | |

| GGUF | IQ2_S | 9.9 | |

| GGUF | IQ2_M | 10.4 | |

| GGUF | Q2_K_S | 10.6 | |

| GGUF | Q2_K | 10.6 | |

| GGUF | IQ3_XXS | 11.3 | lower quality |

| GGUF | IQ3_XS | 11.6 | |

| GGUF | Q3_K_S | 12.2 | |

| GGUF | IQ3_S | 12.2 | |

| GGUF | IQ3_M | 12.4 | |

| GGUF | Q3_K_M | 13.3 | IQ3_S probably better |

| GGUF | Q3_K_L | 13.8 | |

| GGUF | IQ4_XS | 13.9 | |

| GGUF | Q4_0 | 14.5 | |

| GGUF | Q4_K_S | 15.5 | optimal size/speed/quality |

| GGUF | Q4_K_M | 16.8 | fast, recommended |

| GGUF | Q4_1 | 16.0 | |

| GGUF | Q5_K_S | 18.0 | |

| GGUF | Q5_K_M | 19.1 | |

| GGUF | Q6_K | 22.6 | practically like static Q6_K |

Run spiritfather/Gemma4-Gutenberg-26B-A4B-i1-GGUF with guIDE

Download guIDE — the AI-native code editor with local LLM inference and 69 built-in tools.

Download guIDE → · Browse 524k+ models · Compare models

Source: Hugging Face · Compare models