GraySoft
Projects Models Compare Cloud benchmarks FAQ Download guIDE →
Model Intelligence Sheet

Soulfate24/Nanbeige4.2-3B-DSpark-ASHQ1-Remix-GGUF overview

Nanbeige4.2 3B with DSpark ASHQ1 Remix This is a GGUF quantized version of the original model. 📈 Release Benchmarks wiki.test.raw, symmetric FA auto reference…

transformersggufllmnanbeigedsparkquantizationashq1imatrixtext-generationenzhbase_model:Nanbeige/Nanbeige4.2-3Bbase_model:quantized:Nanbeige/Nanbeige4.2-3Blicense:apache-2.0endpoints_compatibleregion:usconversational

Runs locally from ~2.7 MB disk (4 GB VRAM class GPUs with llama.cpp / guIDE).

Downloads
0
Likes
0
Pipeline
text-generation

Repository Files & Downloads

9 GGUF files detected
Direct downloads for local inference
FileTypeQuantizationSizeLink
dspark-ASHQ1-Quality-36pc.ggufGGUFGGUF607.2 MBDownload
imatrix.ggufGGUFGGUF2.7 MBDownload
model-ASHQ1-Compact-33pc.ggufGGUFGGUF2.57 GBDownload
model-ASHQ1-Fidelity-48pc.ggufGGUFGGUF3.73 GBDownload
model-ASHQ1-Mini-30pc.ggufGGUFGGUF2.33 GBDownload
model-ASHQ1-Nano-27pc.ggufGGUFGGUF2.10 GBDownload
model-ASHQ1-Pico-24pc.ggufGGUFGGUF1.87 GBDownload
model-ASHQ1-Precision-42pc.ggufGGUFGGUF3.30 GBDownload
model-ASHQ1-Quality-36pc.ggufGGUFGGUF2.78 GBDownload

Model Details

Model IDSoulfate24/Nanbeige4.2-3B-DSpark-ASHQ1-Remix-GGUF
AuthorSoulfate24
Pipelinetext-generation
Licenseapache-2.0
Base modelNanbeige/Nanbeige4.2-3B,Nanbeige/Nanbeige4.2-3B-DSpark
Last modified2026-09-10T13:29:52.000Z

Model README

---

license: apache-2.0

language:

- en

- zh

library_name: transformers

pipeline_tag: text-generation

base_model:

- Nanbeige/Nanbeige4.2-3B

- Nanbeige/Nanbeige4.2-3B-DSpark

tags:

- llm

- nanbeige

- dspark

- quantization

- gguf

- ashq1

- imatrix

---

Nanbeige4.2-3B with DSpark - ASHQ1-Remix

This is a GGUF quantized version of the original model.

📈 Release Benchmarks (wiki.test.raw, symmetric FA-auto reference)

| Model | Size | PPL | KLD | RMS Δp | top-p | Speed |

| :--- | ---: | ---: | ---: | ---: | ---: | ---: |

| Q8_0 (stock) | 4229 MiB | 34.4742 | 0.0105 | 2.39% | 95.6% | 353 t/s |

| Fidelity-48pc 🥈 | 3824 MiB | 34.3246† | 0.0247 | 3.42% | 94.0% | 371 t/s |

| Precision-42pc ⭐ | 3384 MiB | 34.3791† | 0.0331 | 3.94% | 92.5% | 382 t/s |

| Q6_K-imx (stock) | 3266 MiB | 34.4411† | 0.0335 | 3.99% | 92.4% | 406 t/s |

| Quality-36pc | 2849 MiB | 34.6224 | 0.0743 | 5.80% | 88.3% | 466 t/s |

| Q5_K_M-imx (stock) | 2849 MiB | 34.6224 | 0.0743 | 5.80% | 88.3% | 470 t/s |

| Compact-33pc | 2630 MiB | 34.6636 | 0.1441 | 7.98% | 83.8% | 512 t/s |

| Mini-30pc | 2391 MiB | 33.5214† | 0.1728 | 8.66% | 81.8% | 557 t/s |

| IQ4_XS-imx (stock) | 2268 MiB | 34.8506 | 0.1813 | 9.17% | 81.2% | 457 t/s |

| Nano-27pc ✗ | 2152 MiB | 33.7551† | 0.2191 | 10.12% | 78.6% | 453 t/s |

| IQ3_M-imx (stock) | 1985 MiB | 36.4366 | 0.4492 | 14.31% | 70.5% | 539 t/s |

| Pico-24pc ✗ | 1913 MiB | 37.8642 | 0.3878 | 13.31% | 72.6% | 545 t/s |

ℹ️ About ASHQ1-Remix Suite

Activation-aware GGUF quantization whose every ratio, floor, and cap traces to a measured experiment. Plain-BF16-native first; AutoRound lineage supported with explicit saturation bounds. Full seven-tier ladder validated across six model families.

🔗 Link: https://huggingface.co/Soulfate24/AutoRound-ASHQ1-Remix_Double-Quantization_Suite

Run Soulfate24/Nanbeige4.2-3B-DSpark-ASHQ1-Remix-GGUF with guIDE

Download guIDE — the AI-native code editor with local LLM inference and 69 built-in tools.

Download guIDE → · Browse 524k+ models · Compare models

Source: Hugging Face · Compare models