Model Intelligence Sheet
jamiefutch/GLM-4.6V-MXFP4_MOE-GGUF overview
This is a imatrix MXFP4 MOE quantization of the model GLM 4.6V https://huggingface.co/zai org/GLM 4.6V , based on the imatrix from unsloth. Get the latest llam…
Runs locally from ~948.4 MB disk (4 GB VRAM class GPUs with llama.cpp / guIDE).
Repository Files & Downloads
8 GGUF files detected
Direct downloads for local inference
| File | Type | Quantization | Size | Link |
|---|---|---|---|---|
| GLM-4.6V-MXFP4_MOE-00001-of-00004.gguf | GGUF | GGUF | 14.56 GB | Download |
| GLM-4.6V-MXFP4_MOE-00002-of-00004.gguf | GGUF | GGUF | 14.67 GB | Download |
| GLM-4.6V-MXFP4_MOE-00003-of-00004.gguf | GGUF | GGUF | 14.67 GB | Download |
| GLM-4.6V-MXFP4_MOE-00004-of-00004.gguf | GGUF | GGUF | 12.60 GB | Download |
| mmproj-BF16.gguf | GGUF | BF16 | 1.65 GB | Download |
| mmproj-F16.gguf | GGUF | F16 | 1.60 GB | Download |
| mmproj-F32.gguf | GGUF | F32 | 3.20 GB | Download |
| mmproj-Q8_0.gguf | GGUF | Q8_0 | 948.4 MB | Download |
Model Details
| Model ID | jamiefutch/GLM-4.6V-MXFP4_MOE-GGUF |
|---|---|
| Author | jamiefutch |
| Pipeline | image-text-to-text |
| License | — |
| Base model | zai-org/GLM-4.6V |
| Last modified | 2026-07-06T09:42:50.000Z |
Model README
---
pipeline_tag: image-text-to-text
base_model:
- zai-org/GLM-4.6V
---
This is a imatrix MXFP4_MOE quantization of the model GLM-4.6V, based on the imatrix from unsloth.
Get the latest llama.cpp in order to run it.
For the mmproj, try to use the largest version you can fit in memory in order to get the best results.
F32 > BF16 > F16 > Q8_0.
Run jamiefutch/GLM-4.6V-MXFP4_MOE-GGUF with guIDE
Download guIDE — the AI-native code editor with local LLM inference and 69 built-in tools.
Source: Hugging Face · Compare models