GraySoft
Projects Models Compare Cloud benchmarks FAQ Download guIDE →
Model Intelligence Sheet

constructai/VibeThinker-1.5B-GGUF overview

constructai/VibeThinker 1.5B GGUF This is a quantized version of the original VibeThinker 1.5B , converted to the GGUF format for efficient CPU/GPU inference w…

ggufmathcodereasoninggpqainstruction-followingtext-generationbase_model:WeiboAI/VibeThinker-1.5Bbase_model:quantized:WeiboAI/VibeThinker-1.5Blicense:mitendpoints_compatibleregion:usconversational

Runs locally from ~416.3 MB disk (4 GB VRAM class GPUs with llama.cpp / guIDE).

Downloads
0
Likes
0
Pipeline
text-generation

Repository Files & Downloads

28 GGUF files detected
Direct downloads for local inference
FileTypeQuantizationSizeLink
VibeThinker-1.5B-GGUF-F16.ggufGGUFF162.88 GBDownload
VibeThinker-1.5B-GGUF-Q2_K.ggufGGUFQ2_K645.0 MBDownload
VibeThinker-1.5B-GGUF-Q3_K_M.ggufGGUFQ3_K_M786.0 MBDownload
VibeThinker-1.5B-GGUF-Q3_K_S.ggufGGUFQ3_K_S725.7 MBDownload
VibeThinker-1.5B-GGUF-Q4_K_M.ggufGGUFQ4_K_M940.4 MBDownload
VibeThinker-1.5B-GGUF-Q4_K_S.ggufGGUFQ4_K_S896.8 MBDownload
VibeThinker-1.5B-GGUF-Q5_K_M.ggufGGUFQ5_K_M1.05 GBDownload
VibeThinker-1.5B-GGUF-Q5_K_S.ggufGGUFQ5_K_S1.02 GBDownload
VibeThinker-1.5B-GGUF-Q6_K.ggufGGUFQ6_K1.19 GBDownload
VibeThinker-1.5B-GGUF-Q8_0.ggufGGUFQ8_01.53 GBDownload
VibeThinker-1.5B-GGUF-UD-IQ1_M.ggufGGUFIQ1_M442.9 MBDownload
VibeThinker-1.5B-GGUF-UD-IQ1_S.ggufGGUFIQ1_S416.3 MBDownload
VibeThinker-1.5B-GGUF-UD-IQ2_M.ggufGGUFIQ2_M573.2 MBDownload
VibeThinker-1.5B-GGUF-UD-IQ2_XXS.ggufGGUFIQ2_XXS487.3 MBDownload
VibeThinker-1.5B-GGUF-UD-IQ3_S.ggufGGUFIQ3_S727.1 MBDownload
VibeThinker-1.5B-GGUF-UD-IQ3_XXS.ggufGGUFIQ3_XXS637.8 MBDownload
VibeThinker-1.5B-GGUF-UD-IQ4_NL.ggufGGUFIQ4_NL893.0 MBDownload
VibeThinker-1.5B-GGUF-UD-IQ4_XS.ggufGGUFIQ4_XS854.2 MBDownload
VibeThinker-1.5B-GGUF-UD-Q2_K_XL.ggufGGUFQ2_K_XL645.0 MBDownload
VibeThinker-1.5B-GGUF-UD-Q3_K_M.ggufGGUFQ3_K_M786.0 MBDownload
VibeThinker-1.5B-GGUF-UD-Q3_K_XL.ggufGGUFQ3_K_XL839.4 MBDownload
VibeThinker-1.5B-GGUF-UD-Q4_K_XL.ggufGGUFQ4_K_XL940.4 MBDownload
VibeThinker-1.5B-GGUF-UD-Q5_K_M.ggufGGUFQ5_K_M1.05 GBDownload
VibeThinker-1.5B-GGUF-UD-Q5_K_S.ggufGGUFQ5_K_S1.02 GBDownload
VibeThinker-1.5B-GGUF-UD-Q5_K_XL.ggufGGUFQ5_K_XL1.05 GBDownload
VibeThinker-1.5B-GGUF-UD-Q6_K.ggufGGUFQ6_K1.19 GBDownload
VibeThinker-1.5B-GGUF-UD-Q6_K_XL.ggufGGUFQ6_K_XL1.19 GBDownload
VibeThinker-1.5B-GGUF-UD-Q8_K_XL.ggufGGUFQ8_K_XL1.53 GBDownload

Model Details

Model IDconstructai/VibeThinker-1.5B-GGUF
Authorconstructai
Pipelinetext-generation
Licensemit
Base modelWeiboAI/VibeThinker-1.5B
Last modified2026-06-19T19:41:51.000Z

Model README

---

license: mit

base_model:

  • WeiboAI/VibeThinker-1.5B

pipeline_tag: text-generation

tags:

  • math
  • code
  • reasoning
  • gpqa
  • instruction-following
  • gguf

---

constructai/VibeThinker-1.5B-GGUF

This is a quantized version of the original VibeThinker-1.5B , converted to the GGUF format for efficient CPU/GPU inference with llama.cpp, Ollama, or any GGUF‑compatible runner.

---

Original Model

---

Available Quantizations

Choose the quantization that fits your needs:

| Quantization | File Size |

|--------------|-----------|

| UD-IQ1_S | 437 MB |

| UD-IQ1_M | 464 MB |

| UD-IQ2_XXS | 511 MB |

| Q2_K | 676 MB |

| UD-IQ2_M | 601 MB |

| UD-Q2_K_XL | 676 MB |

| UD-IQ3_XXS | 669 MB |

| Q3_K_S | 761 MB |

| UD-IQ3_S | 762 MB |

| Q3_K_M | 824 MB |

| UD-Q3_K_M | 824 MB |

| UD-Q3_K_XL | 880 MB |

| UD-IQ4_XS | 896 MB |

| Q4_K_S | 940 MB |

| UD-IQ4_NL | 936 MB |

| Q4_K_M | 986 MB |

| UD-Q4_K_XL | 986 MB |

| Q5_K_S | 1.1 GB |

| UD-Q5_K_S | 1.1 GB |

| Q5_K_M | 1.13 GB |

| UD-Q5_K_M | 1.13 GB |

| UD-Q5_K_XL | 1.13 GB |

| Q6_K | 1.27 GB |

| UD-Q6_K | 1.27 GB |

| UD-Q6_K_XL | 1.27 GB |

| Q8_0 | 1.65 GB |

| UD-Q8_K_XL | 1.65 GB |

| F16 | 3.09 GB |

For a 1.5B‑parameter model, even the larger files are quite manageable. Here’s what I recommend: F16 (3.09 GB) or Q8_0 (1.65 GB).

The other quants are also usable!

---

Usage

With ollama

ollama run hf.co/constructai/VibeThinker-1.5B-GGUF:F16

---

With llama.cpp

llama-server -hf constructai/VibeThinker-1.5B-GGUF:VibeThinker-1.5B-GGUF-F16.gguf

or

llama-cli -hf constructai/VibeThinker-1.5B-GGUF:VibeThinker-1.5B-GGUF-F16.gguf

---

With LM Studio

lms get constructai/VibeThinker-1.5B-GGUF@F16

---

Run constructai/VibeThinker-1.5B-GGUF with guIDE

Download guIDE — the AI-native code editor with local LLM inference and 69 built-in tools.

Download guIDE → · Browse 524k+ models · Compare models

Source: Hugging Face · Compare models