GraySoft
Projects Models Compare Cloud benchmarks FAQ Download guIDE →
Model Intelligence Sheet

ggml-org/Qwen3-8B-Base-GGUF overview

Qwen3 8B Base Run with https://llama.app bash llama serve hf ggml org/Qwen3 8B Base GGUF Source models https://huggingface.co/Qwen/Qwen3 8B Base IMPORTANT This…

ggufquantizedtext-generationbase_model:Qwen/Qwen3-8B-Basebase_model:quantized:Qwen/Qwen3-8B-Baselicense:apache-2.0endpoints_compatibleregion:usconversational

Runs locally from ~8.11 GB disk (12 GB VRAM class GPUs with llama.cpp / guIDE).

Downloads
0
Likes
0
Pipeline
text-generation
Author

Repository Files & Downloads

2 GGUF files detected
Direct downloads for local inference
FileTypeQuantizationSizeLink
Qwen3-8B-Base-BF16.ggufGGUFBF1615.26 GBDownload
Qwen3-8B-Base-Q8_0.ggufGGUFQ8_08.11 GBDownload

Model Details

Model IDggml-org/Qwen3-8B-Base-GGUF
Authorggml-org
Pipelinetext-generation
Licenseapache-2.0
Base modelQwen/Qwen3-8B-Base
Last modified2026-07-28T11:13:54.000Z

Model README

---

license: apache-2.0

pipeline_tag: text-generation

tags:

  • gguf
  • quantized

base_model:

  • Qwen/Qwen3-8B-Base

---

Qwen3-8B-Base

Run with https://llama.app

llama serve -hf ggml-org/Qwen3-8B-Base-GGUF

Source models

  • https://huggingface.co/Qwen/Qwen3-8B-Base

> [!IMPORTANT]

> This model is automatically converted using https://github.com/ggml-org/convert

Run ggml-org/Qwen3-8B-Base-GGUF with guIDE

Download guIDE — the AI-native code editor with local LLM inference and 69 built-in tools.

Download guIDE → · Browse 524k+ models · Compare models

Source: Hugging Face · Compare models