EnlistedGhost/Magistral-Small-2509-Vision-GGUF overview
<img src="https://huggingface.co/EnlistedGhost/Magistral Small 2509 Vision GGUF/resolve/main/resources/Magistral Icon MistralAI.png" alt="Magistral image" widt…
Runs locally from ~1.64 GB disk (4 GB VRAM class GPUs with llama.cpp / guIDE).
Repository Files & Downloads
| File | Type | Quantization | Size | Link |
|---|---|---|---|---|
| Magistral-Small-2509-Vision-BF16.gguf | GGUF | BF16 | 43.92 GB | Download |
| Magistral-Small-2509-Vision-Q2_K.gguf | GGUF | Q2_K | 8.59 GB | Download |
| Magistral-Small-2509-Vision-Q2_K_L.gguf | GGUF | Q2_K_L | 10.06 GB | Download |
| Magistral-Small-2509-Vision-Q2_K_M.gguf | GGUF | Q2_K_M | 8.09 GB | Download |
| Magistral-Small-2509-Vision-Q2_K_S.gguf | GGUF | Q2_K_S | 8.42 GB | Download |
| Magistral-Small-2509-Vision-Q3_K_L.gguf | GGUF | Q3_K_L | 11.23 GB | Download |
| Magistral-Small-2509-Vision-Q3_K_M.gguf | GGUF | Q3_K_M | 10.41 GB | Download |
| Magistral-Small-2509-Vision-Q3_K_S.gguf | GGUF | Q3_K_S | 9.93 GB | Download |
| Magistral-Small-2509-Vision-Q3_K_XL.gguf | GGUF | Q3_K_XL | 12.40 GB | Download |
| Magistral-Small-2509-Vision-Q4_K_M.gguf | GGUF | Q4_K_M | 11.76 GB | Download |
| Magistral-Small-2509-Vision-Q4_K_S.gguf | GGUF | Q4_K_S | 12.61 GB | Download |
| Magistral-Small-2509-Vision-Q4_K_XL.gguf | GGUF | Q4_K_XL | 13.81 GB | Download |
| Magistral-Small-2509-Vision-Q5_K_M.gguf | GGUF | Q5_K_M | 15.57 GB | Download |
| Magistral-Small-2509-Vision-Q5_K_S.gguf | GGUF | Q5_K_S | 15.27 GB | Download |
| Magistral-Small-2509-Vision-Q5_K_XL.gguf | GGUF | Q5_K_XL | 14.70 GB | Download |
| Magistral-Small-2509-Vision-Q6_K.gguf | GGUF | Q6_K | 18.02 GB | Download |
| Magistral-Small-2509-Vision-Q6_K_L.gguf | GGUF | Q6_K_L | 25.83 GB | Download |
| Magistral-Small-2509-Vision-Q6_K_M.gguf | GGUF | Q6_K_M | 18.32 GB | Download |
| Magistral-Small-2509-Vision-Q6_K_XL.gguf | GGUF | Q6_K_XL | 19.49 GB | Download |
| Magistral-Small-2509-Vision-Q8_0.gguf | GGUF | Q8_0 | 23.33 GB | Download |
| mmproj-Magistral-Small-2509-Vision-F32.gguf | GGUF | F32 | 1.64 GB | Download |
Model Details
| Model ID | EnlistedGhost/Magistral-Small-2509-Vision-GGUF |
|---|---|
| Author | EnlistedGhost |
| Pipeline | image-text-to-text |
| License | apache-2.0 |
| Base model | EnlistedGhost/Magistral-Small-2509-Vision |
| Last modified | 2026-07-11T14:23:14.000Z |
Model README
---
license: apache-2.0
language:
- en
- ru
- uk
base_model:
- EnlistedGhost/Magistral-Small-2509-Vision
base_model_relation: quantized
new_version: EnlistedGhost/Magistral-Small-2509-Vision-GGUF
pipeline_tag: image-text-to-text
tags:
- Magistral
- Vision
- Conversational
- Multimodal
- Image-Text-to-Text
- GGUF
- GGML
- Quant
- Quantized
- Llama.cpp
- Ollama
- MistralAI
- 24B
---
<img src="https://huggingface.co/EnlistedGhost/Magistral-Small-2509-Vision-GGUF/resolve/main/resources/Magistral_Icon_MistralAI.png" alt="Magistral image" width="395" height="752">
-----------------------------------------------<br /> - Model Details and Specifications: -<br />-----------------------------------------------
-----------------------------------------------<br /> - Update July 11th 2026 -<br />-----------------------------------------------
- Fixed MMPROJ Multi-Modal Vision projector. (Re-Uploaded with working projector)
- Uploaded brand new Q2_K_M, Q3_K_M, Q4_K_M, Q5_K_XL, Q6_K_L quants, using the GH5TS method! These offer significantly higher performance and quality of response over previously uploaded release files with the same or similar size of file!
- All files are being re-Uploaded with brand new and custom Quantize formulae: Over the next couple days the files will be updated with the new "GH5TS" (Pronounced: Ghosts) Quant method! This quantization method is a dynamic quantizaton that bases all quants no matter the size to include Q5_K in layers that are critical to cognitive abilities of the model while allowing for lower Quantization bit values for non-crucial layers.
- New chat template! Yes, a new chat template that completely fixes previous issues for llama.cpp users and retains even higher performance than the originally made template for Ollama/Xllama users as well!
<br /><br />
------------------------------------------------------------
Magistral-Small-2509-Vision (24B Parameters)
------------------------------------------------------------
Description: <br />
This model was re-configured with MistralAI's original Magistral-Small vision tower,
using only official MistralAI files, weights and model data. This has resulted in a
Vision Multimodal version of Magistral-Small-2509 including fully functional Vision (re-enabled). <br /><br />
The Chat-template and System-prompt have been reworked and customized to improve
performance and quality across all Quantizated Files. No modifications,
edits, or additional configurations are required to use this model with Ollama/Llama.cpp
Both Vision and Text work. (^.^) <br />
This release contains: <br />
Llama.cpp, Ollama, and Xllama compatible GGUF converted and Quantized model files <br />
(Compatible with Xllama, Ollama, and Llama.cpp) <br />
IMPORTANT NOTICE as of (GMT-8) 07:00 July 11th 2026: <br />
Please note: The chat-template AND the system-prompt have been rewritten and finalized but differ from what
MistralAI made available. You can still use MistralAI's default template and prompt, however it is recommended to be
using what is provided within this release. Below is a copy of the new chat template (Default System Message Removed for Ease of Reading).
Chat Template:
{%- set ns = namespace(remMessage=false, hasSys=false, injSystem=true) -%}
{%- for msg in messages -%}
{%- if msg.role == "system" -%}
{%- set ns.hasSys = true -%}
{%- endif -%}
{%- endfor -%}
{{- bos_token }}
{%- for msg in messages -%}
{%- if ns.injSystem -%}
[SYSTEM_PROMPT]
{%- if ns.hasSys -%}
{{ msg.content }}
{%- else -%}
'Default System Message Is provided in the actual GGUF file here'
{%- endif -%}
[/SYSTEM_PROMPT]
{%- set ns.injSystem = false -%}
{%- endif -%}
{%- if (messages|length - loop.index0) < 29 -%}
{%- set ns.remMessage = true -%}
{%- endif -%}
{%- if ns.remMessage -%}
{%- if msg.role == "user" -%}
[INST]
{%- if msg.content is string %}
{{ msg.content }}
{%- else %}
{%- for block in msg.content %}
{%- if block.type == 'text' %}
{{- block.text }}
{%- if block.type in ['image', 'image_url'] %}
[IMG]
{%- endif %}
{%- endfor %}
{%- endif %}
[/INST]
{%- elif msg.role == "assistant" -%}
{{ msg.content }}
{%- endif -%}
{%- endif -%}
{{- eos_token }}
{%- endfor -%}
<br /><br />
Happy LLM Inferrencing,<br />
-- Jon Z (EnlistedGhost)
---------------------------------------------------
---------------------------------------------------<br /> - Conversion and GGUF Quantization: -<br />---------------------------------------------------
Quantized GGUF version of:
- EnlistedGhost/Magistral-Small-2509-Vision <br /> (by MistralAI - modified by EnlistedGhost)
Original Model Link (Safetensors):
Converted using: Llama.cpp (b9840)
Quantized using: Llama.cpp (b9760)
---------------------------------------------
--------------------------------------<br /> ---- How to run this Model ---- <br /> --------------------------------------
Compatible Software (Required to use this Model) <br />
You can run this model by using either Ollama (or) Llama.cpp <br />
(Below are instruction on running these GGUF files with Ollama)
How to run this Model using Ollama <br />
You can run this model by using the "ollama run" command.<br />
Simply copy & paste one of the commands from the list below into<br />
your console, terminal or power-shell window.
| Quant Type | File Size | Command |
|:-----------|:----------|:--------|
| QX_X | 0.00 GB | (Currently Uploading Files, Check again very soon!) |
Vision Projector (Files) <br />
mmproj (Vision Projector) Files
| Quant Type | File Size | Download Link |
|:-----------|:----------|:--------|
| Q8_0 | 465 MB | [mmproj Magistral-Small-2509-Vision Projector:Q8_0]|
| F16 | 870 MB | [mmproj Magistral-Small-2509-Vision Projector:F16]|
| F32 | 1.74 GB | [mmproj Magistral-Small-2509-Vision Projector:F32]|
-----------------------------------------------
-------------------------------------------------<br /> - Legal, Citations and Usage Details: -<br />-------------------------------------------------
Intended Use
Same as original:
Out-of-Scope Use
Same as original:
Bias, Risks, and Limitations
Same as original:
Evaluation
- This model has NOT been evaluated in any form, scope or method of use.
- !!! USE AT YOUR OWN RISK !!!
- !!! NO WARRANTY IS PROVIDED OF ANY KIND !!!
---------------------------------------------
Citation (Original Paper)
[MistalAI Magistral-Small-2509 Original Paper]
Detailed Release Information
- Originally Developed by: [MistralAI]
- Modified with Vision re-Enabled by: [EnlistedGhost]
- MMPROJ (Vision) Quantized by: [EnlistedGhost]
- Model Quantized for GGUF by: [EnlistedGhost]
- Model type & format: [Quantized/GGUF]
- License type: [Apache-2.0]
Model Card Authors and Contact
Run EnlistedGhost/Magistral-Small-2509-Vision-GGUF with guIDE
Download guIDE — the AI-native code editor with local LLM inference and 69 built-in tools.
Source: Hugging Face · Compare models