GraySoft
Projects Models Compare Cloud benchmarks FAQ Download guIDE →

xero0000 author hub

Qwen3.6 35B A3B — vram13: an all VRAM ultra quant 12.98 GB, imatrix A 35B MoE that fits entirely in 18 GB of VRAM and decodes at 122–159 t/s on a pair of mid range gaming GPUs RTX 3060 Ti 8 GB + RTX 3080 10 GB , with quality within ~1 % perplexity of Q8 0. Most sub 24 GB rigs ru…

Models
4
Downloads
0
xero0000/Qwen3.6-35B-A3B-vram13-GGUF
0 downloads · text-generation
xero0000/Qwopus3.6-35B-A3B-Coder-vram13-GGUF
0 downloads · text-generation
xero0000/Qwen3.6-27B-vram14-GGUF
0 downloads · text-generation
xero0000/Qwen3.6-35B-A3B-128E-Pruned-GGUF
0 downloads · text-generation

Run models locally with guIDE

Download guIDE — the AI-native code editor with local LLM inference and 69 built-in tools.

Download guIDE → · Browse 524k+ models · Compare models