bogdan-radulescu author hub
gemma 4 26B A4B it | Asymmetric 2 bit Routed Expert Quant A GGUF quantization of gemma 4 26B A4B it built for larger than RAM / SSD streaming inference . Instead of quantizing every tensor to the same width, it quantizes asymmetrically : the routed expert weights ≈89% of the byt…
Models
2
Downloads
0
Run models locally with guIDE
Download guIDE — the AI-native code editor with local LLM inference and 69 built-in tools.