GraySoft
Projects Models Compare Cloud benchmarks FAQ Download guIDE โ†’
Model Intelligence Sheet

Soulfate24/MiniCPM5-2B-DSpark-ASHQ1-Remix-GGUF overview

MiniCPM5 2B with DSpark ASHQ1 Remix This is a GGUF quantized version of the original model. ๐Ÿ“ˆ Release Benchmarks wiki.test.raw, symmetric FA auto reference | โ€ฆ

transformersggufminicpmminicpm5llamatext-generationlong-contexttool-callingon-deviceedge-aidsparkquantizationashq1imatrixenzhbase_model:openbmb/MiniCPM5-2Bbase_model:quantized:openbmb/MiniCPM5-2Blicense:apache-2.0endpoints_compatibleregion:usconversational

Runs locally from ~3.0 MB disk (4 GB VRAM class GPUs with llama.cpp / guIDE).

Downloads
0
Likes
0
Pipeline
text-generation

Repository Files & Downloads

9 GGUF files detected
Direct downloads for local inference
FileTypeQuantizationSizeLink
dspark-ASHQ1-Quality-36pc.ggufGGUFGGUF227.6 MBDownload
imatrix.ggufGGUFGGUF3.0 MBDownload
model-ASHQ1-Compact-33pc.ggufGGUFGGUF1.55 GBDownload
model-ASHQ1-Fidelity-48pc.ggufGGUFGGUF2.26 GBDownload
model-ASHQ1-Mini-30pc.ggufGGUFGGUF1.41 GBDownload
model-ASHQ1-Nano-27pc.ggufGGUFGGUF1.27 GBDownload
model-ASHQ1-Pico-24pc.ggufGGUFGGUF1.13 GBDownload
model-ASHQ1-Precision-42pc.ggufGGUFGGUF1.99 GBDownload
model-ASHQ1-Quality-36pc.ggufGGUFGGUF1.68 GBDownload

Model Details

Model IDSoulfate24/MiniCPM5-2B-DSpark-ASHQ1-Remix-GGUF
AuthorSoulfate24
Pipelinetext-generation
Licenseapache-2.0
Base modelopenbmb/MiniCPM5-2B,openbmb/MiniCPM5-2B-DSpark
Last modified2026-09-10T13:29:56.000Z

Model README

---

license: apache-2.0

language:

- en

- zh

library_name: transformers

pipeline_tag: text-generation

base_model:

- openbmb/MiniCPM5-2B

- openbmb/MiniCPM5-2B-DSpark

tags:

- minicpm

- minicpm5

- llama

- text-generation

- long-context

- tool-calling

- on-device

- edge-ai

- dspark

- quantization

- gguf

- ashq1

- imatrix

---

MiniCPM5-2B with DSpark - ASHQ1-Remix

This is a GGUF quantized version of the original model.

๐Ÿ“ˆ Release Benchmarks (wiki.test.raw, symmetric FA-auto reference)

| Model | Size | PPL | KLD | RMS ฮ”p | top-p | Speed |

| :--- | ---: | ---: | ---: | ---: | ---: | ---: |

| Q8_0 (stock) | 4229 MiB | 34.4742 | 0.0105 | 2.39% | 95.6% | 353 t/s |

| Fidelity-48pc ๐Ÿฅˆ | 3824 MiB | 34.3246โ€  | 0.0247 | 3.42% | 94.0% | 371 t/s |

| Precision-42pc โญ | 3384 MiB | 34.3791โ€  | 0.0331 | 3.94% | 92.5% | 382 t/s |

| Q6_K-imx (stock) | 3266 MiB | 34.4411โ€  | 0.0335 | 3.99% | 92.4% | 406 t/s |

| Quality-36pc | 2849 MiB | 34.6224 | 0.0743 | 5.80% | 88.3% | 466 t/s |

| Q5_K_M-imx (stock) | 2849 MiB | 34.6224 | 0.0743 | 5.80% | 88.3% | 470 t/s |

| Compact-33pc | 2630 MiB | 34.6636 | 0.1441 | 7.98% | 83.8% | 512 t/s |

| Mini-30pc | 2391 MiB | 33.5214โ€  | 0.1728 | 8.66% | 81.8% | 557 t/s |

| IQ4_XS-imx (stock) | 2268 MiB | 34.8506 | 0.1813 | 9.17% | 81.2% | 457 t/s |

| Nano-27pc โœ— | 2152 MiB | 33.7551โ€  | 0.2191 | 10.12% | 78.6% | 453 t/s |

| IQ3_M-imx (stock) | 1985 MiB | 36.4366 | 0.4492 | 14.31% | 70.5% | 539 t/s |

| Pico-24pc โœ— | 1913 MiB | 37.8642 | 0.3878 | 13.31% | 72.6% | 545 t/s |

โ„น๏ธ About ASHQ1-Remix Suite

Activation-aware GGUF quantization whose every ratio, floor, and cap traces to a measured experiment. Plain-BF16-native first; AutoRound lineage supported with explicit saturation bounds. Full seven-tier ladder validated across six model families.

๐Ÿ”— Link: https://huggingface.co/Soulfate24/AutoRound-ASHQ1-Remix_Double-Quantization_Suite

Run Soulfate24/MiniCPM5-2B-DSpark-ASHQ1-Remix-GGUF with guIDE

Download guIDE โ€” the AI-native code editor with local LLM inference and 69 built-in tools.

Download guIDE โ†’ ยท Browse 524k+ models ยท Compare models

Source: Hugging Face ยท Compare models