GraySoft
Projects Models Compare Cloud benchmarks FAQ Download guIDE →
Model Intelligence Sheet

Koshkasa/sophosympatheia_Magistry-24B-v1.1-IQ4_KS.gguf overview

What's that? The goal : Make a medium quality quant of sophosympatheia/Magistry 24B v1.1 using SOTA quant types from ik llama.cpp, allowing the resulting gguf …

ik_llama.cppggufquantizediq4_ks4 bitroleplaymixed precisiontext-generationbase_model:sophosympatheia/Magistry-24B-v1.1base_model:quantized:sophosympatheia/Magistry-24B-v1.1license:otherendpoints_compatibleregion:usimatrixconversational

Runs locally from ~12.66 GB disk (16 GB VRAM class GPUs with llama.cpp / guIDE).

Downloads
0
Likes
0
Pipeline
text-generation
Author

Repository Files & Downloads

1 GGUF files detected
Direct downloads for local inference
FileTypeQuantizationSizeLink
sophosympatheia_Magistry-24B-v1.1-IQ4_KS.ggufGGUFIQ4_KS12.66 GBDownload

Model Details

Model IDKoshkasa/sophosympatheia_Magistry-24B-v1.1-IQ4_KS.gguf
AuthorKoshkasa
Pipelinetext-generation
Licenseother
Base modelsophosympatheia/Magistry-24B-v1.1
Last modified2026-07-27T21:44:29.000Z

Model README

---

license: other

base_model:

  • sophosympatheia/Magistry-24B-v1.1

library_name: ik_llama.cpp

pipeline_tag: text-generation

tags:

  • gguf
  • quantized
  • ik_llama.cpp
  • iq4_ks
  • 4 bit
  • roleplay
  • mixed precision

quantized_by: Koshkasa

base_model_relation: quantized

---

What's that?

The goal: Make a medium quality quant of sophosympatheia/Magistry-24B-v1.1 using SOTA quant types from ik_llama.cpp, allowing the resulting gguf to fit into 16gb VRAM with KVO, accounting for system overhead.

The result: Mixed precision quantization of sophosympatheia/Magistry-24B-v1.1

quantized with ik_llama.cpp build: 9d07d868

incompatible with mainline llama.cpp

Layout

| Layer | Dims | Dims | Quant |

| --- | --- | --- | --- |

| token\_embd | 5120 | 131072.0 | iq4\_k |

| | | | |

| | blk| 40| |

| attn\_k | 5120 | 1024 | iq6\_k |

| attn\_norm | 5120 | 1 | f32 |

| attn\_q | 5120 | 4096 | iq6\_k |

| attn\_v | 5120 | 1024 | iq6\_k |

| ffn\_down | 32768 | 5120 | iq4\_k |

| ffn\_gate | 5120 | 32768 | iq4\_ks |

| ffn\_norm | 5120 | 1 | f32 |

| ffn\_up | 5120 | 32768 | iq4\_ks |

| attn\_output | 4096 | 5120 | iq6\_k |

| | | | |

| output | 5120 | 131072 | iq6\_k |

| output\_norm | 5120 | 1 | f32 |

Cheers

MistralAI - the beloved base model(s).

ikawrakow and contributors of ik_llama.cpp - I probably misused your wonderful creation.

sophosympatheia - for the merge effort.

Everyone whose finetunes were included in the merge!

bartowski - for the imatrix + the myriad of quants we all benefit from.

Run Koshkasa/sophosympatheia_Magistry-24B-v1.1-IQ4_KS.gguf with guIDE

Download guIDE — the AI-native code editor with local LLM inference and 69 built-in tools.

Download guIDE → · Browse 524k+ models · Compare models

Source: Hugging Face · Compare models