GraySoft
Projects Models Compare Cloud benchmarks FAQ Download guIDE →
Model Intelligence Sheet

rpatel622/Mellum2-12B-A2.5B-Thinking-Dflash-GGUF-FIXED overview

Mellum2 12B A2.5B Thinking Dflash GGUF Original draft model: https://huggingface.co/RedHatAI/Mellum2 12B A2.5B Thinking Dflash Original target model: https://h…

ggufbase_model:JetBrains/Mellum2-12B-A2.5B-Thinkingbase_model:quantized:JetBrains/Mellum2-12B-A2.5B-Thinkinglicense:apache-2.0endpoints_compatibleregion:usfeature-extraction

Runs locally from ~515.0 MB disk (4 GB VRAM class GPUs with llama.cpp / guIDE).

Downloads
384
Likes
0
Pipeline
Author

Repository Files & Downloads

5 GGUF files detected
Direct downloads for local inference
FileTypeQuantizationSizeLink
mellum2-dflash.Q4_K_M.ggufGGUFGGUF515.0 MBDownload
mellum2-dflash.Q5_K_S.ggufGGUFGGUF578.6 MBDownload
mellum2-dflash.Q8_0.ggufGGUFGGUF848.0 MBDownload
mellum2-dflash.bf16.ggufGGUFGGUF1.56 GBDownload
mellum2-dflash.f16.ggufGGUFGGUF1.56 GBDownload

Model Details

Model IDrpatel622/Mellum2-12B-A2.5B-Thinking-Dflash-GGUF-FIXED
Authorrpatel622
Pipeline
Licenseapache-2.0
Base modelJetBrains/Mellum2-12B-A2.5B-Thinking
Last modified2026-07-14T16:25:54.000Z

Model README

---

license: apache-2.0

base_model:

  • JetBrains/Mellum2-12B-A2.5B-Thinking

---

Mellum2-12B-A2.5B-Thinking-Dflash GGUF

Original draft model:

https://huggingface.co/RedHatAI/Mellum2-12B-A2.5B-Thinking-Dflash

Original target model:

https://huggingface.co/JetBrains/Mellum2-12B-A2.5B-Thinking

Converted using:

https://github.com/Anbeeld/beellama.cpp

This repository only contains GGUF conversions.

Run rpatel622/Mellum2-12B-A2.5B-Thinking-Dflash-GGUF-FIXED with guIDE

Download guIDE — the AI-native code editor with local LLM inference and 69 built-in tools.

Download guIDE → · Browse 524k+ models · Compare models

Source: Hugging Face · Compare models