snuggguf/Qwen3-4B-Instruct-2507-GGUF overview
🧣 SnugGGUF: Qwen3 4B Instruct 2507 Every model. Every quant. Always free. Always tested. ⚡ Quick Start Ollama: ollama run snuggguf/qwen3 4b instruct 2507 📦 A…
Runs locally from ~2.33 GB disk (4 GB VRAM class GPUs with llama.cpp / guIDE).
Repository Files & Downloads
| File | Type | Quantization | Size | Link |
|---|---|---|---|---|
| Qwen3-4B-Instruct-2507-Q4_K_M.gguf | GGUF | Q4_K_M | 2.33 GB | Download |
Model Details
| Model ID | snuggguf/Qwen3-4B-Instruct-2507-GGUF |
|---|---|
| Author | snuggguf |
| Pipeline | text-generation |
| License | apache-2.0 |
| Base model | Qwen/Qwen3-4B-Instruct-2507 |
| Last modified | 2026-08-24T16:28:38.000Z |
Model README
---
license: apache-2.0
base_model: Qwen/Qwen3-4B-Instruct-2507
tags: [gguf, quantized, text-generation, snuggguf]
---
🧣 SnugGGUF: Qwen3-4B-Instruct-2507
> Every model. Every quant. Always free. Always tested.
⚡ Quick Start
Ollama:
>ollama run snuggguf/qwen3-4b-instruct-2507
📦 Available Quants
| Quant | Size | Tested |
|-------|------|--------|
| Q4_K_M | ~2.33 GB | ✅ Verified locally with Ollama |
More quants (Q5_K_M, Q6_K) coming soon.
✅ Verification
- GGUF magic bytes verified during conversion ✅
- File size integrity confirmed ✅
- Local Ollama inference test passed ✅
- Runs 100% offline — no internet required ✅
🐾 About SnugGGUF
SnugGGUF quantizes every major open model — Qwen, DeepSeek, Mistral, Llama, Gemma, GLM, Nemotron, Kimi, gpt-oss and more.
Every file is inference-tested before going public. No broken quants. Ever.
Run snuggguf/Qwen3-4B-Instruct-2507-GGUF with guIDE
Download guIDE — the AI-native code editor with local LLM inference and 69 built-in tools.
Source: Hugging Face · Compare models