GraySoft
Projects Models Compare Cloud benchmarks FAQ Download guIDE →
Model Intelligence Sheet

Jerome0207/Huihui-GLM-5.2-abliterated-UD-Q2_K_MXFP4-GGUF overview

Huihui GLM 5.2 Abliterated — UD Q2 K MXFP4 GGUF This is a single quant operational mirror for RunPod cached models. It contains only the seven UD Q2 K MXFP4 GG…

ggufglmglm-5.2abliteratedquantizedllama.cpptext-generationenzhbase_model:zai-org/GLM-5.2base_model:quantized:zai-org/GLM-5.2license:mitendpoints_compatibleregion:usimatrixconversational

Runs locally from ~9.0 MB disk (4 GB VRAM class GPUs with llama.cpp / guIDE).

Downloads
0
Likes
0
Pipeline
text-generation

Repository Files & Downloads

7 GGUF files detected
Direct downloads for local inference
FileTypeQuantizationSizeLink
GLM-5.2-UD-Q2_K_MXFP4-00001-of-00007.ggufGGUFQ2_K_MXFP49.0 MBDownload
GLM-5.2-UD-Q2_K_MXFP4-00002-of-00007.ggufGGUFQ2_K_MXFP445.71 GBDownload
GLM-5.2-UD-Q2_K_MXFP4-00003-of-00007.ggufGGUFQ2_K_MXFP445.45 GBDownload
GLM-5.2-UD-Q2_K_MXFP4-00004-of-00007.ggufGGUFQ2_K_MXFP445.45 GBDownload
GLM-5.2-UD-Q2_K_MXFP4-00005-of-00007.ggufGGUFQ2_K_MXFP445.45 GBDownload
GLM-5.2-UD-Q2_K_MXFP4-00006-of-00007.ggufGGUFQ2_K_MXFP446.29 GBDownload
GLM-5.2-UD-Q2_K_MXFP4-00007-of-00007.ggufGGUFQ2_K_MXFP46.90 GBDownload

Model Details

Model IDJerome0207/Huihui-GLM-5.2-abliterated-UD-Q2_K_MXFP4-GGUF
AuthorJerome0207
Pipelinetext-generation
Licensemit
Base modelzai-org/GLM-5.2
Last modified2026-07-29T07:14:29.000Z

Model README

---

license: mit

language:

- en

- zh

pipeline_tag: text-generation

base_model:

- zai-org/GLM-5.2

tags:

- gguf

- glm

- glm-5.2

- abliterated

- quantized

- llama.cpp

---

Huihui GLM-5.2 Abliterated — UD-Q2_K_MXFP4 GGUF

This is a single-quant operational mirror for RunPod cached models. It contains

only the seven UD-Q2_K_MXFP4 GGUF shards from the pinned huihui-ai source.

It does not claim authorship of the model or its quantization.

Destination repository: Jerome0207/Huihui-GLM-5.2-abliterated-UD-Q2_K_MXFP4-GGUF

Provenance

  1. zai-org/GLM-5.2
  2. huihui-ai abliterated GGUF

Source revision:

994200a058539c553f0977002e58a2f05de845c7.

The seven GGUF shards total 252,600,253,568 bytes. See

checksums.sha256 for the source SHA-256 object identifiers.

Runtime

The repository is intended for a current CUDA build of

llama.cpp. Start with the first shard;

llama.cpp discovers the remaining split files automatically:

GLM-5.2-UD-Q2_K_MXFP4-00001-of-00007.gguf

The validated deployment target uses three 96 GB GPUs, 131,072 context tokens,

one parallel slot and Q8_0 K/V caches.

License and limitations

The source model card declares the MIT license. See NOTICE for attribution

and exact provenance.

This is an “abliterated” derivative with intentionally reduced refusal

behavior. It is not a safety-tuned public service. Deploy it only behind

authentication and rate limits, and run coding/security agents in isolated

sandboxes without production credentials.

GLM-5.2 support in llama.cpp is evolving. Full DSA, Lightning Indexer,

IndexShare and MTP behavior must be validated before treating this deployment

as equivalent to the official inference implementation.

Run Jerome0207/Huihui-GLM-5.2-abliterated-UD-Q2_K_MXFP4-GGUF with guIDE

Download guIDE — the AI-native code editor with local LLM inference and 69 built-in tools.

Download guIDE → · Browse 524k+ models · Compare models

Source: Hugging Face · Compare models