agentionai author hub
Qwen3.8 27B DFlash2 draft model, ROCmFP4 FAST GGUF A 4.25 bpw ROCmFP4 requantisation of z lab/Qwen3.8 27B DFlash2 https://huggingface.co/z lab/Qwen3.8 27B DFlash2 , for use as a speculative decoding sidecar with a Qwen3.8 27B target. Measured on an AMD Strix Halo Radeon 8060S , …
Models
7
Downloads
52,408
Run models locally with guIDE
Download guIDE — the AI-native code editor with local LLM inference and 69 built-in tools.