GraySoft
Projects Models Compare Cloud benchmarks FAQ Download guIDE →
Model Intelligence Sheet

ggml-org/LongCat-Flash-Chat-GGUF overview

| Quant | Size | Mixture | PPL | 1 Mean PPL Q /PPL base | KLD | | : | : | : | : | : | : | | Q8 0 | 557.39 GiB 8.51 BPW | Q8 0 | 3.275749 ± 0.017863 | +0.2918% …

ggufbase_model:meituan-longcat/LongCat-Flash-Chatbase_model:quantized:meituan-longcat/LongCat-Flash-Chatendpoints_compatibleregion:usimatrixconversational

Runs locally from ~5.2 MB disk (4 GB VRAM class GPUs with llama.cpp / guIDE).

Downloads
151
Likes
1
Pipeline
Author

Repository Files & Downloads

5 GGUF files detected
Direct downloads for local inference
FileTypeQuantizationSizeLink
IQ1_S/LongCat-Flash-Chat-IQ1_S-00001-of-00004.ggufGGUFIQ1_S5.2 MBDownload
IQ1_S/LongCat-Flash-Chat-IQ1_S-00002-of-00004.ggufGGUFIQ1_S46.41 GBDownload
IQ1_S/LongCat-Flash-Chat-IQ1_S-00003-of-00004.ggufGGUFIQ1_S46.36 GBDownload
IQ1_S/LongCat-Flash-Chat-IQ1_S-00004-of-00004.ggufGGUFIQ1_S13.41 GBDownload
imatrix.ggufGGUFGGUF805.1 MBDownload

Model Details

Model IDggml-org/LongCat-Flash-Chat-GGUF
Authorggml-org
Pipeline
License
Base modelmeituan-longcat/LongCat-Flash-Chat
Last modified2026-08-09T02:46:37.000Z

Model README

---

base_model:

  • meituan-longcat/LongCat-Flash-Chat

---

| Quant | Size | Mixture | PPL | 1-(Mean PPL(Q)/PPL(base)) | KLD |

| :---- | :-------------------- | :------ | :------------------- | :------------------------ | :------------------ |

| Q8_0 | 557.39 GiB (8.51 BPW) | Q8_0 | 3.275749 ± 0.017863 | +0.2918% | 0.003764 ± 0.000045 |

| IQ1_S | 106.19 GiB (1.62 BPW) | IQ1_S | 11.010332 ± 0.077291 | +237.0971% | 1.331246 ± 0.004350 |

!kld_graph

!ppl_graph

Run ggml-org/LongCat-Flash-Chat-GGUF with guIDE

Download guIDE — the AI-native code editor with local LLM inference and 69 built-in tools.

Download guIDE → · Browse 524k+ models · Compare models

Source: Hugging Face · Compare models