Mojo24x7 author hub
Qwen3.6 35B A3B — NPU aware GGUF for Rockchip RK3588 A GGUF quantisation of Qwen3.6 35B A3B built specifically for the RK3588 NPU , using the RKNPU2 backend in rk llama.cpp . Runs on a 16 GB Radxa ROCK 5B+ : 20.4 tok/s prefill, 4.8 tok/s decode on a 1967 token prompt — the faste…
Models
2
Downloads
0
Run models locally with guIDE
Download guIDE — the AI-native code editor with local LLM inference and 69 built-in tools.