GraySoft
Projects Models Compare Cloud benchmarks FAQ Download guIDE →
Model Intelligence Sheet

Iloqt/Versi-StyleTune-31B-GGUF overview

license: apache 2.0 base model: Iloqt/Versi StyleTune 31B base model relation: quantized tags: gguf gemma 4 31B merge mergekit reasoning creative writing rolep…

ggufgemma-431Bmergemergekitreasoningcreative writingroleplayconversationalbase_model:Iloqt/Versi-StyleTune-31Bbase_model:quantized:Iloqt/Versi-StyleTune-31Blicense:apache-2.0endpoints_compatibleregion:us

Runs locally from ~11.53 GB disk (12 GB VRAM class GPUs with llama.cpp / guIDE).

Downloads
687
Likes
0
Pipeline
Author

Repository Files & Downloads

15 GGUF files detected
Direct downloads for local inference
FileTypeQuantizationSizeLink
Versi-StyleTune-31B-Q2_K.ggufGGUFQ2_K11.53 GBDownload
Versi-StyleTune-31B-Q3_K_M.ggufGGUFQ3_K_M14.80 GBDownload
Versi-StyleTune-31B-Q4_K_M-attn8-HB.ggufGGUFQ4_K_M25.39 GBDownload
Versi-StyleTune-31B-Q4_K_M-hb16.ggufGGUFQ4_K_M21.58 GBDownload
Versi-StyleTune-31B-Q4_K_M.ggufGGUFQ4_K_M18.14 GBDownload
Versi-StyleTune-31B-Q5_K_M-attn8-HB.ggufGGUFQ5_K_M27.41 GBDownload
Versi-StyleTune-31B-Q5_K_M-hb16.ggufGGUFQ5_K_M24.52 GBDownload
Versi-StyleTune-31B-Q5_K_M.ggufGGUFQ5_K_M21.25 GBDownload
Versi-StyleTune-31B-Q6_K-attn8-HB.ggufGGUFQ6_K29.56 GBDownload
Versi-StyleTune-31B-Q6_K-hb16.ggufGGUFQ6_K27.64 GBDownload
Versi-StyleTune-31B-Q6_K.ggufGGUFQ6_K24.55 GBDownload
Versi-StyleTune-31B-Q8_0.ggufGGUFQ8_031.79 GBDownload
Versi-StyleTune-31B.i1-Q4_K_M-hb16.ggufGGUFQ4_K_M21.58 GBDownload
Versi-StyleTune-31B.i1-Q6_K-hb16.ggufGGUFQ6_K27.64 GBDownload
Versi-StyleTune-BF16.ggufGGUFBF1659.82 GBDownload

Model Details

Model IDIloqt/Versi-StyleTune-31B-GGUF
AuthorIloqt
Pipeline
Licenseapache-2.0
Base modelIloqt/Versi-StyleTune-31B
Last modified2026-07-17T12:25:50.000Z

Model README

---

license: apache-2.0

base_model:

  • Iloqt/Versi-StyleTune-31B

base_model_relation: quantized

tags:

- gguf

- gemma-4

- 31B

- merge

- mergekit

- reasoning

- creative writing

- roleplay

- conversational

not-for-all-audiences: true

---

Versi-StyleTune-31B (GGUF)

GGUF quants of Iloqt/Versi-StyleTune-head: a text-only Gemma 4 31B with Nimbz/Versipellis-31B as the base and the lm_head (output projection) grafted from Gryphe/Gemma-4-31B-StyleTune.

All credits go to Nimbz and Gryphe for the original models, I only committed the merge.

Variants

Standard quants:

  • Q2_K, Q3_K_M, Q4_K_M, Q5_K_M, Q6_K, Q8_0 — body-only quantization.

hb16 variants (head + embeddings kept at BF16):

  • Q4_K_M-hb16, Q5_K_M-hb16, Q6_K-hb16
  • Preserves the grafted StyleTune lm_head and embed_tokens at full precision while quantizing the rest of the body to K-quant.

attn8-HB variants (Q8_0 attention + BF16 head + embeddings):

  • Q4_K_M-attn8-HB, Q5_K_M-attn8-HB, Q6_K-attn8-HB
  • Adds Q8_0 attention layers on top of the hb16 protection.

Notes

  • Run with the Gemma 4 chat template; thinking off by default.
  • Like all Gemma 4 models, benefits from repetition penalty or DRY to avoid token loops.

Run Iloqt/Versi-StyleTune-31B-GGUF with guIDE

Download guIDE — the AI-native code editor with local LLM inference and 69 built-in tools.

Download guIDE → · Browse 524k+ models · Compare models

Source: Hugging Face · Compare models