GraySoft
Projects Models Compare Cloud benchmarks FAQ Download guIDE →

felippeburk author hub

Qwen3.8 27B NVFP4 MTP GGUF NVFP4 quantized Qwen3.8 27B with MTP multi token prediction draft layers, packaged as a single GGUF for llama.cpp speculative decoding on Blackwell GPUs RTX 5090 / 5080 . What this is Quantization: NVFP4 weights + KV cache , converted to GGUF via llama…

Models
2
Downloads
2,231
felippeburk/Qwen3.8-27B-NVFP4-MTP-GGUF
2,231 downloads · text-generation

Run models locally with guIDE

Download guIDE — the AI-native code editor with local LLM inference and 69 built-in tools.

Download guIDE → · Browse 524k+ models · Compare models