cygnal author hub
Qwen3.8 Flash Next Uncensored — IQ4XS NGQ4 GGUF — AMD Strix Halo gfx1151 First working GGUF build of Qwen3.8 Flash Next qwen4exp architecture with vision, running on stock llama.cpp — no custom tensor formats or forked runtime required. Built from orcarouter/Qwen3.8 Flash Next U…
Models
2
Downloads
3,852
Run models locally with guIDE
Download guIDE — the AI-native code editor with local LLM inference and 69 built-in tools.