felippeburk author hub
Qwen3.8 27B NVFP4 MTP GGUF NVFP4 quantized Qwen3.8 27B with MTP multi token prediction draft layers, packaged as a single GGUF for llama.cpp speculative decoding on Blackwell GPUs RTX 5090 / 5080 . What this is Quantization: NVFP4 weights + KV cache , converted to GGUF via llama…
Models
2
Downloads
2,231
Run models locally with guIDE
Download guIDE — the AI-native code editor with local LLM inference and 69 built-in tools.