spiritfather/Gemma4-Gutenberg-26B-A4B-GGUF overview
About static quants of https://huggingface.co/nbeerbower/Gemma4 Gutenberg 26B A4B Note: Gemma4 Gutenberg 26B A4B https://huggingface.co/nbeerbower/Gemma4 Guten…
Runs locally from ~9.86 GB disk (12 GB VRAM class GPUs with llama.cpp / guIDE).
Repository Files & Downloads
| File | Type | Quantization | Size | Link |
|---|---|---|---|---|
| Gemma4-Gutenberg-26B-A4B.IQ4_XS.gguf | GGUF | GGUF | 13.10 GB | Download |
| Gemma4-Gutenberg-26B-A4B.Q2_K.gguf | GGUF | GGUF | 9.86 GB | Download |
| Gemma4-Gutenberg-26B-A4B.Q3_K_L.gguf | GGUF | GGUF | 12.88 GB | Download |
| Gemma4-Gutenberg-26B-A4B.Q3_K_M.gguf | GGUF | GGUF | 12.37 GB | Download |
| Gemma4-Gutenberg-26B-A4B.Q3_K_S.gguf | GGUF | GGUF | 11.38 GB | Download |
| Gemma4-Gutenberg-26B-A4B.Q4_K_M.gguf | GGUF | GGUF | 15.64 GB | Download |
| Gemma4-Gutenberg-26B-A4B.Q4_K_S.gguf | GGUF | GGUF | 14.40 GB | Download |
| Gemma4-Gutenberg-26B-A4B.Q5_K_M.gguf | GGUF | GGUF | 17.82 GB | Download |
| Gemma4-Gutenberg-26B-A4B.Q5_K_S.gguf | GGUF | GGUF | 16.75 GB | Download |
| Gemma4-Gutenberg-26B-A4B.Q6_K.gguf | GGUF | GGUF | 21.08 GB | Download |
| Gemma4-Gutenberg-26B-A4B.Q8_0.gguf | GGUF | GGUF | 25.02 GB | Download |
Model Details
| Model ID | spiritfather/Gemma4-Gutenberg-26B-A4B-GGUF |
|---|---|
| Author | spiritfather |
| Pipeline | text-generation |
| License | gemma |
| Base model | nbeerbower/Gemma4-Gutenberg-26B-A4B |
| Last modified | 2026-07-07T15:12:48.000Z |
Model README
---
base_model: nbeerbower/Gemma4-Gutenberg-26B-A4B
language:
- en
library_name: transformers
license: gemma
quantized_by: spiritfather
tags:
- text-generation
- creative-writing
- gguf
---
About
static quants of https://huggingface.co/nbeerbower/Gemma4-Gutenberg-26B-A4B
> Note: Gemma4-Gutenberg-26B-A4B is published as a LoRA adapter, not a full model. These GGUFs are that adapter merged onto its base google/gemma-4-26B-A4B-it (the text LM extracted from the multimodal base), then quantized.
<!-- provided-files -->
Usage
If you are unsure how to use GGUF files, refer to one of TheBloke's READMEs for more details, including on how to concatenate multi-part files.
Provided Quants
(sorted by size, not necessarily quality. IQ-quants are often preferable over similar sized non-IQ quants)
| Link | Type | Size/GB | Notes |
|:-----|:-----|--------:|:------|
| GGUF | Q2_K | 10.6 | |
| GGUF | Q3_K_S | 12.2 | |
| GGUF | Q3_K_M | 13.3 | lower quality |
| GGUF | Q3_K_L | 13.8 | |
| GGUF | IQ4_XS | 14.1 | |
| GGUF | Q4_K_S | 15.5 | fast, recommended |
| GGUF | Q4_K_M | 16.8 | fast, recommended |
| GGUF | Q5_K_S | 18.0 | |
| GGUF | Q5_K_M | 19.1 | |
| GGUF | Q6_K | 22.6 | very good quality |
| GGUF | Q8_0 | 26.9 | fast, best quality |
Run spiritfather/Gemma4-Gutenberg-26B-A4B-GGUF with guIDE
Download guIDE — the AI-native code editor with local LLM inference and 69 built-in tools.
Source: Hugging Face · Compare models