GraySoft
Projects Models Compare Cloud benchmarks FAQ Download guIDE →
Model Intelligence Sheet

kaitchup/MiniMax-M3-GGUF-MoQ overview

<div align="center" <img src="https://cdn uploads.huggingface.co/production/uploads/64b93e6bd6c468ac7536607e/mj6xac74jHGLqymiovObc.png" alt="The Kaitchup AI on…

ggufbase_model:MiniMaxAI/MiniMax-M3base_model:quantized:MiniMaxAI/MiniMax-M3license:otherendpoints_compatibleregion:usimatrixconversational

Runs locally from ~123.18 GB disk (32 GB+ VRAM class GPUs with llama.cpp / guIDE).

Downloads
1,228
Likes
7
Pipeline
Author

Repository Files & Downloads

8 GGUF files detected
Direct downloads for local inference
FileTypeQuantizationSizeLink
MoQ-2.5.ggufGGUFGGUF123.18 GBDownload
MoQ-2.75.ggufGGUFGGUF135.20 GBDownload
MoQ-3.0.ggufGGUFGGUF148.57 GBDownload
MoQ-3.25.ggufGGUFGGUF153.24 GBDownload
MoQ-3.5.ggufGGUFGGUF171.27 GBDownload
MoQ-3.75.ggufGGUFGGUF185.59 GBDownload
MoQ-4.75.ggufGGUFGGUF189.31 GBDownload
MoQ-5.0.ggufGGUFGGUF239.41 GBDownload

Model Details

Model IDkaitchup/MiniMax-M3-GGUF-MoQ
Authorkaitchup
Pipeline
Licenseother
Base modelMiniMaxAI/MiniMax-M3
Last modified2026-06-30T20:21:28.000Z

Model README

---

license: other

license_name: minimax-community

license_link: LICENSE

base_model:

  • MiniMaxAI/MiniMax-M3

---

<div align="center">

<img

src="https://cdn-uploads.huggingface.co/production/uploads/64b93e6bd6c468ac7536607e/mj6xac74jHGLqymiovObc.png"

alt="The Kaitchup -- AI on a Budget"

style="width: 100%; max-width: 100%; height: auto; display: inline-block; margin-bottom: 0.5em; margin-top: 0.5em;"

/>

<div style="display: flex; justify-content: center; gap: 0.5em; margin-bottom: 1em;">

<a href="https://kaitchup.substack.com/subscribe"><strong>Subscribe and Support</strong></a>

</div>

</div>

GGUF models made with the method ("Mixture of Quantizations") proposed by Waleed Ahmad.

I also used Unsloth M3's imatrix for calibration.

More details and evaluation here:

MiniMax M3 GGUF Quantization: From 852 GB to ~150 GB Without Breaking Accuracy

!image

Avoid using the MoQ-2.5.

  • Compute Sponsorship: Verda. I used 2 B300s for quantization and evaluation.

Run kaitchup/MiniMax-M3-GGUF-MoQ with guIDE

Download guIDE — the AI-native code editor with local LLM inference and 69 built-in tools.

Download guIDE → · Browse 524k+ models · Compare models

Source: Hugging Face · Compare models