liodon-ai/MaralGPT-Mythos-9B-2606-GGUF-imatrix-GGUF overview
MaralGPT Mythos 9B 2606 GGUF — iMatrix GGUF GGUF quantizations of MaralGPT/MaralGPT Mythos 9B 2606 GGUF https://huggingface.co/MaralGPT/MaralGPT Mythos 9B 2606…
Runs locally from ~3.36 GB disk (4 GB VRAM class GPUs with llama.cpp / guIDE).
Repository Files & Downloads
| File | Type | Quantization | Size | Link |
|---|---|---|---|---|
| MaralGPT-Mythos-9B-2606-GGUF-IQ2_M.gguf | GGUF | IQ2_M | 3.36 GB | Download |
| MaralGPT-Mythos-9B-2606-GGUF-IQ3_M.gguf | GGUF | IQ3_M | 4.11 GB | Download |
| MaralGPT-Mythos-9B-2606-GGUF-IQ4_XS.gguf | GGUF | IQ4_XS | 4.84 GB | Download |
| MaralGPT-Mythos-9B-2606-GGUF-Q4_K_M.gguf | GGUF | Q4_K_M | 5.24 GB | Download |
| MaralGPT-Mythos-9B-2606-GGUF-Q5_K_M.gguf | GGUF | Q5_K_M | 6.02 GB | Download |
| MaralGPT-Mythos-9B-2606-GGUF-Q6_K.gguf | GGUF | Q6_K | 6.85 GB | Download |
| MaralGPT-Mythos-9B-2606-GGUF-Q8_0.gguf | GGUF | Q8_0 | 8.87 GB | Download |
Model Details
| Model ID | liodon-ai/MaralGPT-Mythos-9B-2606-GGUF-imatrix-GGUF |
|---|---|
| Author | liodon-ai |
| Pipeline | text-generation |
| License | other |
| Base model | MaralGPT/MaralGPT-Mythos-9B-2606-GGUF |
| Last modified | 2026-07-13T04:54:45.000Z |
Model README
---
license: other
base_model: MaralGPT/MaralGPT-Mythos-9B-2606-GGUF
base_model_relation: quantized
pipeline_tag: text-generation
library_name: gguf
tags:
- gguf
- ollama
- local-llm
- llama.cpp
- lm-studio
- quantized
- imatrix
- sub-4-bit
- qwen3.5
quantized_by: liodon-ai
---
MaralGPT-Mythos-9B-2606-GGUF — iMatrix GGUF
GGUF quantizations of MaralGPT/MaralGPT-Mythos-9B-2606-GGUF, published by Liodon AI.
Quick Start
llama.cpp
llama-cli -hf liodon-ai/MaralGPT-Mythos-9B-2606-GGUF-imatrix-GGUF:Q4_K_M
Ollama
ollama run hf.co/liodon-ai/MaralGPT-Mythos-9B-2606-GGUF-imatrix-GGUF:Q4_K_M
LM Studio / Jan — search liodon-ai/MaralGPT-Mythos-9B-2606-GGUF-imatrix-GGUF and pick your quant.
Quants
| Quant | Size | VRAM est. | Notes |
|-------|------|-----------|-------|
| IQ2_M | 3.61 GB | ~4 GB | 2-bit, iMatrix — smallest usable |
| IQ3_M | 4.42 GB | ~5 GB | 3-bit, iMatrix — great quality/size tradeoff |
| IQ4_XS | 5.20 GB | ~6 GB | 4-bit extra-small, iMatrix |
| Q4_K_M | 5.63 GB | ~6 GB | 4-bit, iMatrix-calibrated (recommended) |
| Q5_K_M | 6.47 GB | ~7 GB | 5-bit, iMatrix-calibrated |
| Q6_K | 7.36 GB | ~8 GB | 6-bit, iMatrix-calibrated, near-lossless |
| Q8_0 | 9.53 GB | ~11 GB | 8-bit, essentially lossless |
What is iMatrix?
Standard quantization treats all weights equally. iMatrix runs 128 calibration chunks through
the full-precision model to find which weights matter most, then allocates more precision where
it counts. At Q2/Q3/Q4 this means noticeably better coherence and instruction-following —
same file size, better output.
Calibration: 2M tokens of WikiText-103.
> Also see plain (non-iMatrix) quants: liodon-ai/MaralGPT-Mythos-9B-2606-GGUF-GGUF
Source
- Model: MaralGPT/MaralGPT-Mythos-9B-2606-GGUF
- License: other
---
Quantized by Liodon AI
Run liodon-ai/MaralGPT-Mythos-9B-2606-GGUF-imatrix-GGUF with guIDE
Download guIDE — the AI-native code editor with local LLM inference and 69 built-in tools.
Source: Hugging Face · Compare models