GraySoft
Projects Models Compare Cloud benchmarks FAQ Download guIDE →
Model Intelligence Sheet

Michionlion/Nanbeige4.2-3B-GGUF-WebGPU overview

Nanbeige 4.2 3B Q4 K M for browser WebGPU This is a six part, browser friendly GGUF of Nanbeige/Nanbeige4.2 3B https://huggingface.co/Nanbeige/Nanbeige4.2 3B .…

ggufnanbeigewebgpuwllamabase_model:Nanbeige/Nanbeige4.2-3Bbase_model:quantized:Nanbeige/Nanbeige4.2-3Blicense:apache-2.0endpoints_compatibleregion:usconversational

Runs locally from ~190.6 MB disk (4 GB VRAM class GPUs with llama.cpp / guIDE).

Downloads
0
Likes
0
Pipeline

Repository Files & Downloads

6 GGUF files detected
Direct downloads for local inference
FileTypeQuantizationSizeLink
Nanbeige4.2-3B-Q4_K_M-00001-of-00006.ggufGGUFQ4_K_M403.1 MBDownload
Nanbeige4.2-3B-Q4_K_M-00002-of-00006.ggufGGUFQ4_K_M468.8 MBDownload
Nanbeige4.2-3B-Q4_K_M-00003-of-00006.ggufGGUFQ4_K_M472.9 MBDownload
Nanbeige4.2-3B-Q4_K_M-00004-of-00006.ggufGGUFQ4_K_M459.2 MBDownload
Nanbeige4.2-3B-Q4_K_M-00005-of-00006.ggufGGUFQ4_K_M460.9 MBDownload
Nanbeige4.2-3B-Q4_K_M-00006-of-00006.ggufGGUFQ4_K_M190.6 MBDownload

Model Details

Model IDMichionlion/Nanbeige4.2-3B-GGUF-WebGPU
AuthorMichionlion
Pipeline
Licenseapache-2.0
Base modelNanbeige/Nanbeige4.2-3B
Last modified2026-07-27T18:46:09.000Z

Model README

---

base_model: Nanbeige/Nanbeige4.2-3B

license: apache-2.0

library_name: gguf

tags:

- nanbeige

- gguf

- webgpu

- wllama

---

Nanbeige 4.2 3B Q4_K_M for browser WebGPU

This is a six-part, browser-friendly GGUF of

Nanbeige/Nanbeige4.2-3B.

Each shard is below 500 MiB so it can be downloaded and cached by wllama.

The Q4_K_M tensor payload comes from

Tdamre/Nanbeige4.2-3B-GGUF.

The tensors were not requantized. The container metadata was normalized from

the Nanbeige fork's older convention to the convention merged into upstream

llama.cpp:

  • nanbeige.block_count: 22 physical layers
  • nanbeige.num_loops: 2
  • nanbeige.skip_loop_final_norm: false

Validated with llama.cpp merge commit

b77d646751d01c0962bc203b6809e9d94f7d50b7.

Load the first shard; llama.cpp/wllama discovers the other five from their

standard split names.

The canonical public assets are the

nanbeige4.2-3b-q4km-v1 GitHub release.

SHA-256

8fdb05799b34cfd3d3b11afaf22f0cf17bbb26a04558ea887115dd1569d93d3c  Nanbeige4.2-3B-Q4_K_M-00001-of-00006.gguf
806cdd41859ce1f4956efcd46d1e171accd8c96496b3168d1a82418cac0f3a9b  Nanbeige4.2-3B-Q4_K_M-00002-of-00006.gguf
19e28679b716217bddbf91684cfb23c03f43777065a2ef5cd63519f8e1db9551  Nanbeige4.2-3B-Q4_K_M-00003-of-00006.gguf
4e2f164fedb13384a4ba654f9987fa1c2f9c1b782340c3746caa0c42ef187aa1  Nanbeige4.2-3B-Q4_K_M-00004-of-00006.gguf
146237aade6e3b32eaefc791c10ca7e4b6baaf4768147c856b59f732b2d6370f  Nanbeige4.2-3B-Q4_K_M-00005-of-00006.gguf
d13a2d60af1eb0e61092bd69fcdd48db6a8ff0374f1d96f9b7f7ba1185b37aeb  Nanbeige4.2-3B-Q4_K_M-00006-of-00006.gguf

Run Michionlion/Nanbeige4.2-3B-GGUF-WebGPU with guIDE

Download guIDE — the AI-native code editor with local LLM inference and 69 built-in tools.

Download guIDE → · Browse 524k+ models · Compare models

Source: Hugging Face · Compare models