GraySoft
Projects Models Compare Cloud benchmarks FAQ Download guIDE →
Model Intelligence Sheet

noctrex/Ling-3.0-flash-heretic-MXFP4_MOE-GGUF overview

This is a MXFP4 MOE quantization of the model trohrbaugh / Ling 3.0 flash heretic https://huggingface.co/trohrbaugh/Ling 3.0 flash heretic Quick Start 1. Downl…

gguftext-generationbase_model:trohrbaugh/Ling-3.0-flash-hereticbase_model:quantized:trohrbaugh/Ling-3.0-flash-hereticendpoints_compatibleregion:usconversational

Runs locally from ~13.64 GB disk (16 GB VRAM class GPUs with llama.cpp / guIDE).

Downloads
0
Likes
0
Pipeline
text-generation
Author

Repository Files & Downloads

4 GGUF files detected
Direct downloads for local inference
FileTypeQuantizationSizeLink
Ling-3.0-flash-heretic-MXFP4_MOE-00001-of-00004.ggufGGUFGGUF17.44 GBDownload
Ling-3.0-flash-heretic-MXFP4_MOE-00002-of-00004.ggufGGUFGGUF17.43 GBDownload
Ling-3.0-flash-heretic-MXFP4_MOE-00003-of-00004.ggufGGUFGGUF17.44 GBDownload
Ling-3.0-flash-heretic-MXFP4_MOE-00004-of-00004.ggufGGUFGGUF13.64 GBDownload

Model Details

Model IDnoctrex/Ling-3.0-flash-heretic-MXFP4_MOE-GGUF
Authornoctrex
Pipelinetext-generation
License
Base modeltrohrbaugh/Ling-3.0-flash-heretic
Last modified2026-08-17T20:58:27.000Z

Model README

---

pipeline_tag: text-generation

base_model:

  • trohrbaugh/Ling-3.0-flash-heretic

---

This is a MXFP4_MOE quantization of the model trohrbaugh / Ling-3.0-flash-heretic

Quick Start

  1. Download the latest release of llama.cpp.

Recommended parameters from inclusionAI:

  • temperature=0.6
  • top_p=0.95
  • top_k=20

Performance

| Metric | This model | Original model |

| :----- | :--------: | :---------------------------: |

| KL divergence | 0.0526 | 0 (by definition) |

| Refusals | 0/100 | 87/100 |

Run noctrex/Ling-3.0-flash-heretic-MXFP4_MOE-GGUF with guIDE

Download guIDE — the AI-native code editor with local LLM inference and 69 built-in tools.

Download guIDE → · Browse 524k+ models · Compare models

Source: Hugging Face · Compare models