GraySoft
Projects Models Compare Cloud benchmarks FAQ Download guIDE →
Model Intelligence Sheet

noctrex/Ling-3.0-flash-MXFP4_MOE-GGUF overview

This is a MXFP4 MOE quantization of the model inclusionAI / Ling 3.0 flash https://huggingface.co/inclusionAI/Ling 3.0 flash Quick Start 1. Download the latest…

gguftext-generationbase_model:inclusionAI/Ling-3.0-flashbase_model:quantized:inclusionAI/Ling-3.0-flashendpoints_compatibleregion:usconversational

Runs locally from ~13.64 GB disk (16 GB VRAM class GPUs with llama.cpp / guIDE).

Downloads
0
Likes
1
Pipeline
text-generation
Author

Repository Files & Downloads

4 GGUF files detected
Direct downloads for local inference
FileTypeQuantizationSizeLink
Ling-3.0-flash-MXFP4_MOE-00001-of-00004.ggufGGUFGGUF17.44 GBDownload
Ling-3.0-flash-MXFP4_MOE-00002-of-00004.ggufGGUFGGUF17.43 GBDownload
Ling-3.0-flash-MXFP4_MOE-00003-of-00004.ggufGGUFGGUF17.44 GBDownload
Ling-3.0-flash-MXFP4_MOE-00004-of-00004.ggufGGUFGGUF13.64 GBDownload

Model Details

Model IDnoctrex/Ling-3.0-flash-MXFP4_MOE-GGUF
Authornoctrex
Pipelinetext-generation
License
Base modelinclusionAI/Ling-3.0-flash
Last modified2026-08-17T20:40:33.000Z

Model README

---

pipeline_tag: text-generation

base_model:

  • inclusionAI/Ling-3.0-flash

---

This is a MXFP4_MOE quantization of the model inclusionAI / Ling-3.0-flash

Quick Start

  1. Download the latest release of llama.cpp.

Recommended parameters from inclusionAI:

  • temperature=0.6
  • top_p=0.95
  • top_k=20

Run noctrex/Ling-3.0-flash-MXFP4_MOE-GGUF with guIDE

Download guIDE — the AI-native code editor with local LLM inference and 69 built-in tools.

Download guIDE → · Browse 524k+ models · Compare models

Source: Hugging Face · Compare models