GraySoft
Projects Models Compare Cloud benchmarks FAQ Download guIDE →
Model Intelligence Sheet

ggml-org/Qwen3.8-Flash-Next-GGUF overview

Qwen3.8 Flash Next Run with https://llama.app bash llama serve hf ggml org/Qwen3.8 Flash Next GGUF Source models https://huggingface.co/Qwen/Qwen3.8 Flash Next…

ggufquantizedimage-text-to-textbase_model:Qwen/Qwen3.8-Flash-Nextbase_model:quantized:Qwen/Qwen3.8-Flash-Nextlicense:otherendpoints_compatibleregion:usconversational

Runs locally from ~10.4 MB disk (4 GB VRAM class GPUs with llama.cpp / guIDE).

Downloads
0
Likes
4
Pipeline
image-text-to-text
Author

Repository Files & Downloads

3 GGUF files detected
Direct downloads for local inference
FileTypeQuantizationSizeLink
Qwen3.8-Flash-Next-Q8_0-00001-of-00002.ggufGGUFQ8_010.4 MBDownload
Qwen3.8-Flash-Next-Q8_0-00002-of-00002.ggufGGUFQ8_0151.45 GBDownload
mmproj-Qwen3.8-Flash-Next-Q8_0.ggufGGUFQ8_0588.1 MBDownload

Model Details

Model IDggml-org/Qwen3.8-Flash-Next-GGUF
Authorggml-org
Pipelineimage-text-to-text
Licenseother
Base modelQwen/Qwen3.8-Flash-Next
Last modified2026-08-27T18:06:18.000Z

Model README

---

license: other

pipeline_tag: image-text-to-text

tags:

  • gguf
  • quantized

base_model:

  • Qwen/Qwen3.8-Flash-Next

---

Qwen3.8-Flash-Next

Run with https://llama.app

llama serve -hf ggml-org/Qwen3.8-Flash-Next-GGUF

Source models

  • https://huggingface.co/Qwen/Qwen3.8-Flash-Next

TODOs

  • add info

> [!IMPORTANT]

> This model is automatically converted using https://github.com/ggml-org/convert

Run ggml-org/Qwen3.8-Flash-Next-GGUF with guIDE

Download guIDE — the AI-native code editor with local LLM inference and 69 built-in tools.

Download guIDE → · Browse 524k+ models · Compare models

Source: Hugging Face · Compare models