GraySoft
Projects Models Compare Cloud benchmarks FAQ Download guIDE →
Model Intelligence Sheet

DevQuasar/amd.Instella-MoE-16B-A3B-Think-GGUF overview

<img src="https://raw.githubusercontent.com/csabakecskemeti/devquasar/main/dq logo black transparent.png" width="200"/ https://devquasar.com 'Make knowledge fr…

gguftext-generationbase_model:amd/Instella-MoE-16B-A3B-Thinkbase_model:quantized:amd/Instella-MoE-16B-A3B-Thinkendpoints_compatibleregion:usconversational

Runs locally from ~6.07 GB disk (8 GB VRAM class GPUs with llama.cpp / guIDE).

Downloads
0
Likes
1
Pipeline
text-generation
Author

Repository Files & Downloads

6 GGUF files detected
Direct downloads for local inference
FileTypeQuantizationSizeLink
Q2_K/amd.Instella-MoE-16B-A3B-Think.f16.gguf.Q2_K.ggufGGUFQ2_K6.07 GBDownload
Q3_K_M/amd.Instella-MoE-16B-A3B-Think.f16.gguf.Q3_K_M.ggufGGUFQ3_K_M7.65 GBDownload
Q4_K_M/amd.Instella-MoE-16B-A3B-Think.f16.gguf.Q4_K_M.ggufGGUFQ4_K_M9.75 GBDownload
Q5_K_M/amd.Instella-MoE-16B-A3B-Think.f16.gguf.Q5_K_M.ggufGGUFQ5_K_M11.15 GBDownload
Q6_K/amd.Instella-MoE-16B-A3B-Think.f16.gguf.Q6_K.ggufGGUFQ6_K13.23 GBDownload
Q8_0/amd.Instella-MoE-16B-A3B-Think.f16.gguf.Q8_0.ggufGGUFQ8_015.71 GBDownload

Model Details

Model IDDevQuasar/amd.Instella-MoE-16B-A3B-Think-GGUF
AuthorDevQuasar
Pipelinetext-generation
License
Base modelamd/Instella-MoE-16B-A3B-Think
Last modified2026-07-31T23:01:23.000Z

Model README

---

base_model:

  • amd/Instella-MoE-16B-A3B-Think

pipeline_tag: text-generation

---

<img src="https://raw.githubusercontent.com/csabakecskemeti/devquasar/main/dq_logo_black-transparent.png" width="200"/>

'Make knowledge free for everyone'

Experimental

Use this llama.cpp branch: https://github.com/csabakecskemeti/llama.cpp/tree/instella-moe

Quantized version of: amd/Instella-MoE-16B-A3B-Think

<a href='https://ko-fi.com/L4L416YX7C' target='_blank'><img height='36' style='border:0px;height:36px;' src='https://storage.ko-fi.com/cdn/kofi6.png?v=6' border='0' alt='Buy Me a Coffee at ko-fi.com' /></a>

Run DevQuasar/amd.Instella-MoE-16B-A3B-Think-GGUF with guIDE

Download guIDE — the AI-native code editor with local LLM inference and 69 built-in tools.

Download guIDE → · Browse 524k+ models · Compare models

Source: Hugging Face · Compare models