anik-jha author hub
Qwen3.6 35B A3B coding specialist, 50% experts pruned GGUF Qwen3.6 35B A3B with half of its experts removed 128 of 256 per layer, REAP scoring on a coding heavy calibration mix , quantized with an importance matrix. This is the main artifact of the paper Half the Experts, All th…
Models
2
Downloads
0
Run models locally with guIDE
Download guIDE — the AI-native code editor with local LLM inference and 69 built-in tools.