constructai/VibeThinker-1.5B-GGUF overview
constructai/VibeThinker 1.5B GGUF This is a quantized version of the original VibeThinker 1.5B , converted to the GGUF format for efficient CPU/GPU inference w…
Runs locally from ~416.3 MB disk (4 GB VRAM class GPUs with llama.cpp / guIDE).
Repository Files & Downloads
| File | Type | Quantization | Size | Link |
|---|---|---|---|---|
| VibeThinker-1.5B-GGUF-F16.gguf | GGUF | F16 | 2.88 GB | Download |
| VibeThinker-1.5B-GGUF-Q2_K.gguf | GGUF | Q2_K | 645.0 MB | Download |
| VibeThinker-1.5B-GGUF-Q3_K_M.gguf | GGUF | Q3_K_M | 786.0 MB | Download |
| VibeThinker-1.5B-GGUF-Q3_K_S.gguf | GGUF | Q3_K_S | 725.7 MB | Download |
| VibeThinker-1.5B-GGUF-Q4_K_M.gguf | GGUF | Q4_K_M | 940.4 MB | Download |
| VibeThinker-1.5B-GGUF-Q4_K_S.gguf | GGUF | Q4_K_S | 896.8 MB | Download |
| VibeThinker-1.5B-GGUF-Q5_K_M.gguf | GGUF | Q5_K_M | 1.05 GB | Download |
| VibeThinker-1.5B-GGUF-Q5_K_S.gguf | GGUF | Q5_K_S | 1.02 GB | Download |
| VibeThinker-1.5B-GGUF-Q6_K.gguf | GGUF | Q6_K | 1.19 GB | Download |
| VibeThinker-1.5B-GGUF-Q8_0.gguf | GGUF | Q8_0 | 1.53 GB | Download |
| VibeThinker-1.5B-GGUF-UD-IQ1_M.gguf | GGUF | IQ1_M | 442.9 MB | Download |
| VibeThinker-1.5B-GGUF-UD-IQ1_S.gguf | GGUF | IQ1_S | 416.3 MB | Download |
| VibeThinker-1.5B-GGUF-UD-IQ2_M.gguf | GGUF | IQ2_M | 573.2 MB | Download |
| VibeThinker-1.5B-GGUF-UD-IQ2_XXS.gguf | GGUF | IQ2_XXS | 487.3 MB | Download |
| VibeThinker-1.5B-GGUF-UD-IQ3_S.gguf | GGUF | IQ3_S | 727.1 MB | Download |
| VibeThinker-1.5B-GGUF-UD-IQ3_XXS.gguf | GGUF | IQ3_XXS | 637.8 MB | Download |
| VibeThinker-1.5B-GGUF-UD-IQ4_NL.gguf | GGUF | IQ4_NL | 893.0 MB | Download |
| VibeThinker-1.5B-GGUF-UD-IQ4_XS.gguf | GGUF | IQ4_XS | 854.2 MB | Download |
| VibeThinker-1.5B-GGUF-UD-Q2_K_XL.gguf | GGUF | Q2_K_XL | 645.0 MB | Download |
| VibeThinker-1.5B-GGUF-UD-Q3_K_M.gguf | GGUF | Q3_K_M | 786.0 MB | Download |
| VibeThinker-1.5B-GGUF-UD-Q3_K_XL.gguf | GGUF | Q3_K_XL | 839.4 MB | Download |
| VibeThinker-1.5B-GGUF-UD-Q4_K_XL.gguf | GGUF | Q4_K_XL | 940.4 MB | Download |
| VibeThinker-1.5B-GGUF-UD-Q5_K_M.gguf | GGUF | Q5_K_M | 1.05 GB | Download |
| VibeThinker-1.5B-GGUF-UD-Q5_K_S.gguf | GGUF | Q5_K_S | 1.02 GB | Download |
| VibeThinker-1.5B-GGUF-UD-Q5_K_XL.gguf | GGUF | Q5_K_XL | 1.05 GB | Download |
| VibeThinker-1.5B-GGUF-UD-Q6_K.gguf | GGUF | Q6_K | 1.19 GB | Download |
| VibeThinker-1.5B-GGUF-UD-Q6_K_XL.gguf | GGUF | Q6_K_XL | 1.19 GB | Download |
| VibeThinker-1.5B-GGUF-UD-Q8_K_XL.gguf | GGUF | Q8_K_XL | 1.53 GB | Download |
Model Details
| Model ID | constructai/VibeThinker-1.5B-GGUF |
|---|---|
| Author | constructai |
| Pipeline | text-generation |
| License | mit |
| Base model | WeiboAI/VibeThinker-1.5B |
| Last modified | 2026-06-19T19:41:51.000Z |
Model README
---
license: mit
base_model:
- WeiboAI/VibeThinker-1.5B
pipeline_tag: text-generation
tags:
- math
- code
- reasoning
- gpqa
- instruction-following
- gguf
---
constructai/VibeThinker-1.5B-GGUF
This is a quantized version of the original VibeThinker-1.5B , converted to the GGUF format for efficient CPU/GPU inference with llama.cpp, Ollama, or any GGUF‑compatible runner.
---
Original Model
- Author(s): WeiboAI
- Source: VibeThinker-1.5B
- Original License: MIT
---
Available Quantizations
Choose the quantization that fits your needs:
| Quantization | File Size |
|--------------|-----------|
| UD-IQ1_S | 437 MB |
| UD-IQ1_M | 464 MB |
| UD-IQ2_XXS | 511 MB |
| Q2_K | 676 MB |
| UD-IQ2_M | 601 MB |
| UD-Q2_K_XL | 676 MB |
| UD-IQ3_XXS | 669 MB |
| Q3_K_S | 761 MB |
| UD-IQ3_S | 762 MB |
| Q3_K_M | 824 MB |
| UD-Q3_K_M | 824 MB |
| UD-Q3_K_XL | 880 MB |
| UD-IQ4_XS | 896 MB |
| Q4_K_S | 940 MB |
| UD-IQ4_NL | 936 MB |
| Q4_K_M | 986 MB |
| UD-Q4_K_XL | 986 MB |
| Q5_K_S | 1.1 GB |
| UD-Q5_K_S | 1.1 GB |
| Q5_K_M | 1.13 GB |
| UD-Q5_K_M | 1.13 GB |
| UD-Q5_K_XL | 1.13 GB |
| Q6_K | 1.27 GB |
| UD-Q6_K | 1.27 GB |
| UD-Q6_K_XL | 1.27 GB |
| Q8_0 | 1.65 GB |
| UD-Q8_K_XL | 1.65 GB |
| F16 | 3.09 GB |
For a 1.5B‑parameter model, even the larger files are quite manageable. Here’s what I recommend: F16 (3.09 GB) or Q8_0 (1.65 GB).
The other quants are also usable!
---
Usage
With ollama
ollama run hf.co/constructai/VibeThinker-1.5B-GGUF:F16
---
With llama.cpp
llama-server -hf constructai/VibeThinker-1.5B-GGUF:VibeThinker-1.5B-GGUF-F16.gguf
or
llama-cli -hf constructai/VibeThinker-1.5B-GGUF:VibeThinker-1.5B-GGUF-F16.gguf
---
With LM Studio
lms get constructai/VibeThinker-1.5B-GGUF@F16
---
Run constructai/VibeThinker-1.5B-GGUF with guIDE
Download guIDE — the AI-native code editor with local LLM inference and 69 built-in tools.
Source: Hugging Face · Compare models