GraySoft
Projects Models Compare Cloud benchmarks FAQ Download guIDE →
Model Intelligence Sheet

GCSA-AiLab/Qwen3.8-Flash-Next-FP8-Abliterated-GGUF overview

Qwen3.8 Flash Next FP8 Abliterated GGUF Unofficial community derivative of Qwen/Qwen3.8 Flash Next https://huggingface.co/Qwen/Qwen3.8 Flash Next , published b…

llama.cppggufqwenqwen3.8abliterixquantizedmtpspeculative-decodingmultimodalvision-languageimage-text-to-textmultilingualbase_model:Qwen/Qwen3.8-Flash-Nextbase_model:quantized:Qwen/Qwen3.8-Flash-Nextlicense:otherendpoints_compatibleregion:usconversational

Runs locally from ~2.60 GB disk (4 GB VRAM class GPUs with llama.cpp / guIDE).

Downloads
3,338
Likes
3
Pipeline
image-text-to-text

Repository Files & Downloads

6 GGUF files detected
Direct downloads for local inference
FileTypeQuantizationSizeLink
Q4_KM/Qwen3.8-Flash-Next-Abliterix-Direct-Q4_K_M.ggufGGUFQ4_KM110.95 GBDownload
Q4_KM/mtp-Qwen3.8-Flash-Next-Abliterix-Direct-Q4_K_M.ggufGGUFQ4_KM2.60 GBDownload
Q6_K/Qwen3.8-Flash-Next-Abliterix-Direct-Q6_K.ggufGGUFQ6_K156.11 GBDownload
Q6_K/mtp-Qwen3.8-Flash-Next-Abliterix-Direct-Q6_K.ggufGGUFQ6_K3.18 GBDownload
Q8/Qwen3.8-Flash-Next-Abliterix-Direct-Q8_0.ggufGGUFQ8175.28 GBDownload
Q8/mtp-Qwen3.8-Flash-Next-Abliterix-Direct-Q8_0.ggufGGUFQ83.86 GBDownload

Model Details

Model IDGCSA-AiLab/Qwen3.8-Flash-Next-FP8-Abliterated-GGUF
AuthorGCSA-AiLab
Pipelineimage-text-to-text
Licenseother
Base modelQwen/Qwen3.8-Flash-Next
Last modified2026-10-09T12:13:38.000Z

Model README

---

license: other

license_name: qwen-community-license-1.0

license_link: https://huggingface.co/Qwen/Qwen3.8-Flash-Next/blob/main/LICENSE

base_model:

  • Qwen/Qwen3.8-Flash-Next

pipeline_tag: image-text-to-text

library_name: llama.cpp

tags:

  • qwen
  • qwen3.8
  • gguf
  • abliterix
  • quantized
  • mtp
  • speculative-decoding
  • multimodal
  • vision-language
  • llama.cpp

language:

  • multilingual

---

Qwen3.8-Flash-Next-FP8-Abliterated-GGUF

Unofficial community derivative of Qwen/Qwen3.8-Flash-Next, published by GCSA-AiLab.

这是 GCSA-AiLab 发布的 Qwen/Qwen3.8-Flash-Next 非官方社区衍生版本。

Multimodal Mixture-of-Experts derivative with optional MTP.

多模态混合专家衍生模型,可选 MTP。

Introduction and usage / 介绍与使用

English | 简体中文 | 繁體中文 | 日本語 | 한국어 | Français | Português

Downloads / 下载

Q4_KM · Q6_K · Q8

Technical details · English | 技术细节 · 简体中文

Experimental changes to refusal behavior may alter safeguards. Validate outputs and safety measures before use; follow applicable laws and the Qwen Community License 1.0 upstream terms.

拒答相关行为的实验性调整可能改变安全保护。使用前请验证输出及安全措施,并遵守适用法律和上游许可证。

Run GCSA-AiLab/Qwen3.8-Flash-Next-FP8-Abliterated-GGUF with guIDE

Download guIDE — the AI-native code editor with local LLM inference and 69 built-in tools.

Download guIDE → · Browse 524k+ models · Compare models

Source: Hugging Face · Compare models