GraySoft
Projects Models Compare Cloud benchmarks FAQ Download guIDE →

Mojo24x7 author hub

Qwen3.6 35B A3B — NPU aware GGUF for Rockchip RK3588 A GGUF quantisation of Qwen3.6 35B A3B built specifically for the RK3588 NPU , using the RKNPU2 backend in rk llama.cpp . Runs on a 16 GB Radxa ROCK 5B+ : 20.4 tok/s prefill, 4.8 tok/s decode on a 1967 token prompt — the faste…

Models
2
Downloads
0
Mojo24x7/Qwen3.6-35B-A3B-npuaware-rk3588-GGUF
0 downloads · text-generation
Mojo24x7/Qwen3-30B-A3B-npuaware-rk3588-GGUF
0 downloads · text-generation

Run models locally with guIDE

Download guIDE — the AI-native code editor with local LLM inference and 69 built-in tools.

Download guIDE → · Browse 524k+ models · Compare models