GraySoft
Projects Models Compare Cloud benchmarks FAQ Download guIDE →
Model Intelligence Sheet

mradermacher/alduin-4b-it-base-i1-GGUF overview

About < quantize version: 2 < output tensor quantised: 1 < convert type: hf < vocab type: < tags: nicoboss < quants: Q2 K IQ3 M Q4 K S IQ3 XXS Q3 K M small IQ4…

transformersggufunslothhereticuncensoreddecensoredabliteratedenbase_model:mdamir97/alduin-4b-it-basebase_model:quantized:mdamir97/alduin-4b-it-baselicense:gemmaendpoints_compatibleregion:usimatrixconversational

Runs locally from ~3.3 MB disk (4 GB VRAM class GPUs with llama.cpp / guIDE).

Downloads
2,941
Likes
0
Pipeline

Repository Files & Downloads

25 GGUF files detected
Direct downloads for local inference
FileTypeQuantizationSizeLink
alduin-4b-it-base.i1-IQ1_M.ggufGGUFIQ1_M1.12 GBDownload
alduin-4b-it-base.i1-IQ1_S.ggufGGUFIQ1_S1.06 GBDownload
alduin-4b-it-base.i1-IQ2_M.ggufGGUFIQ2_M1.44 GBDownload
alduin-4b-it-base.i1-IQ2_S.ggufGGUFIQ2_S1.36 GBDownload
alduin-4b-it-base.i1-IQ2_XS.ggufGGUFIQ2_XS1.32 GBDownload
alduin-4b-it-base.i1-IQ2_XXS.ggufGGUFIQ2_XXS1.23 GBDownload
alduin-4b-it-base.i1-IQ3_M.ggufGGUFIQ3_M1.86 GBDownload
alduin-4b-it-base.i1-IQ3_S.ggufGGUFIQ3_S1.81 GBDownload
alduin-4b-it-base.i1-IQ3_XS.ggufGGUFIQ3_XS1.74 GBDownload
alduin-4b-it-base.i1-IQ3_XXS.ggufGGUFIQ3_XXS1.58 GBDownload
alduin-4b-it-base.i1-IQ4_NL.ggufGGUFIQ4_NL2.21 GBDownload
alduin-4b-it-base.i1-IQ4_XS.ggufGGUFIQ4_XS2.12 GBDownload
alduin-4b-it-base.i1-Q2_K.ggufGGUFQ2_K1.62 GBDownload
alduin-4b-it-base.i1-Q2_K_S.ggufGGUFQ2_K_S1.53 GBDownload
alduin-4b-it-base.i1-Q3_K_L.ggufGGUFQ3_K_L2.09 GBDownload
alduin-4b-it-base.i1-Q3_K_M.ggufGGUFQ3_K_M1.96 GBDownload
alduin-4b-it-base.i1-Q3_K_S.ggufGGUFQ3_K_S1.81 GBDownload
alduin-4b-it-base.i1-Q4_0.ggufGGUFQ4_02.21 GBDownload
alduin-4b-it-base.i1-Q4_1.ggufGGUFQ4_12.40 GBDownload
alduin-4b-it-base.i1-Q4_K_M.ggufGGUFQ4_K_M2.33 GBDownload
alduin-4b-it-base.i1-Q4_K_S.ggufGGUFQ4_K_S2.22 GBDownload
alduin-4b-it-base.i1-Q5_K_M.ggufGGUFQ5_K_M2.64 GBDownload
alduin-4b-it-base.i1-Q5_K_S.ggufGGUFQ5_K_S2.58 GBDownload
alduin-4b-it-base.i1-Q6_K.ggufGGUFQ6_K2.98 GBDownload
alduin-4b-it-base.imatrix.ggufGGUFGGUF3.3 MBDownload

Model Details

Model IDmradermacher/alduin-4b-it-base-i1-GGUF
Authormradermacher
Pipeline
Licensegemma
Base modelmdamir97/alduin-4b-it-base
Last modified2026-07-03T09:59:00.000Z

Model README

---

base_model: mdamir97/alduin-4b-it-base

extra_gated_button_content: Acknowledge license

extra_gated_heading: Access Gemma on Hugging Face

extra_gated_prompt: To access Gemma on Hugging Face, you’re required to review and

agree to Google’s usage license. To do this, please ensure you’re logged in to Hugging

Face and click below. Requests are processed immediately.

language:

  • en

library_name: transformers

license: gemma

mradermacher:

readme_rev: 1

quantized_by: mradermacher

tags:

  • unsloth
  • heretic
  • uncensored
  • decensored
  • abliterated

---

About

<!-- ### quantize_version: 2 -->

<!-- ### output_tensor_quantised: 1 -->

<!-- ### convert_type: hf -->

<!-- ### vocab_type: -->

<!-- ### tags: nicoboss -->

<!-- ### quants: Q2_K IQ3_M Q4_K_S IQ3_XXS Q3_K_M small-IQ4_NL Q4_K_M IQ2_M Q6_K IQ4_XS Q2_K_S IQ1_M Q3_K_S IQ2_XXS Q3_K_L IQ2_XS Q5_K_S IQ2_S IQ1_S Q5_K_M Q4_0 IQ3_XS Q4_1 IQ3_S -->

<!-- ### quants_skip: -->

<!-- ### skip_mmproj: -->

weighted/imatrix quants of https://huggingface.co/mdamir97/alduin-4b-it-base

<!-- provided-files -->

For a convenient overview and download list, visit our model page for this model.

static quants are available at https://huggingface.co/mradermacher/alduin-4b-it-base-GGUF

This is a vision model - mmproj files (if any) will be in the static repository.

Usage

If you are unsure how to use GGUF files, refer to one of [TheBloke's

READMEs](https://huggingface.co/TheBloke/KafkaLM-70B-German-V0.1-GGUF) for

more details, including on how to concatenate multi-part files.

Provided Quants

(sorted by size, not necessarily quality. IQ-quants are often preferable over similar sized non-IQ quants)

| Link | Type | Size/GB | Notes |

|:-----|:-----|--------:|:------|

| GGUF | imatrix | 0.1 | imatrix file (for creating your own quants) |

| GGUF | i1-IQ1_S | 1.2 | for the desperate |

| GGUF | i1-IQ1_M | 1.3 | mostly desperate |

| GGUF | i1-IQ2_XXS | 1.4 | |

| GGUF | i1-IQ2_XS | 1.5 | |

| GGUF | i1-IQ2_S | 1.6 | |

| GGUF | i1-IQ2_M | 1.6 | |

| GGUF | i1-Q2_K_S | 1.7 | very low quality |

| GGUF | i1-IQ3_XXS | 1.8 | lower quality |

| GGUF | i1-Q2_K | 1.8 | IQ3_XXS probably better |

| GGUF | i1-IQ3_XS | 2.0 | |

| GGUF | i1-IQ3_S | 2.0 | beats Q3_K* |

| GGUF | i1-Q3_K_S | 2.0 | IQ3_XS probably better |

| GGUF | i1-IQ3_M | 2.1 | |

| GGUF | i1-Q3_K_M | 2.2 | IQ3_S probably better |

| GGUF | i1-Q3_K_L | 2.3 | IQ3_M probably better |

| GGUF | i1-IQ4_XS | 2.4 | |

| GGUF | i1-IQ4_NL | 2.5 | prefer IQ4_XS |

| GGUF | i1-Q4_0 | 2.5 | fast, low quality |

| GGUF | i1-Q4_K_S | 2.5 | optimal size/speed/quality |

| GGUF | i1-Q4_K_M | 2.6 | fast, recommended |

| GGUF | i1-Q4_1 | 2.7 | |

| GGUF | i1-Q5_K_S | 2.9 | |

| GGUF | i1-Q5_K_M | 2.9 | |

| GGUF | i1-Q6_K | 3.3 | practically like static Q6_K |

Here is a handy graph by ikawrakow comparing some lower-quality quant

types (lower is better):

!image.png

And here are Artefact2's thoughts on the matter:

https://gist.github.com/Artefact2/b5f810600771265fc1e39442288e8ec9

FAQ / Model Request

See https://huggingface.co/mradermacher/model_requests for some answers to

questions you might have and/or if you want some other model quantized.

Thanks

I thank my company, nethype GmbH, for letting

me use its servers and providing upgrades to my workstation to enable

this work in my free time. Additional thanks to @nicoboss for giving me access to his private supercomputer, enabling me to provide many more imatrix quants, at much higher quality, than I would otherwise be able to.

<!-- end -->

Run mradermacher/alduin-4b-it-base-i1-GGUF with guIDE

Download guIDE — the AI-native code editor with local LLM inference and 69 built-in tools.

Download guIDE → · Browse 524k+ models · Compare models

Source: Hugging Face · Compare models