GraySoft
Projects Models Compare Cloud benchmarks FAQ Download guIDE →
Model Intelligence Sheet

gghfexp/Kimi-K2.7-Code-GGUF overview

imatrix Quantization of moonshotai/Kimi K2.7 ik llama.cpp quants of moonshotai/Kimi K2.7 using Unsloth's imatrix and Ubergarm's quant recipes . embedding and o…

ggufmlaimatrixconversationalik_llama.cpptext-generationbase_model:moonshotai/Kimi-K2.7-Codebase_model:quantized:moonshotai/Kimi-K2.7-Codelicense:otherendpoints_compatibleregion:us

Runs locally from ~6.6 MB disk (4 GB VRAM class GPUs with llama.cpp / guIDE).

Downloads
74
Likes
0
Pipeline
text-generation
Author

Repository Files & Downloads

56 GGUF files detected
Direct downloads for local inference
FileTypeQuantizationSizeLink
IQ2_KL/Kimi-K2.7-Code-IQ2_KL-00001-of-00014.ggufGGUFIQ2_KL6.6 MBDownload
IQ2_KL/Kimi-K2.7-Code-IQ2_KL-00002-of-00014.ggufGGUFIQ2_KL26.51 GBDownload
IQ2_KL/Kimi-K2.7-Code-IQ2_KL-00003-of-00014.ggufGGUFIQ2_KL27.27 GBDownload
IQ2_KL/Kimi-K2.7-Code-IQ2_KL-00004-of-00014.ggufGGUFIQ2_KL27.27 GBDownload
IQ2_KL/Kimi-K2.7-Code-IQ2_KL-00005-of-00014.ggufGGUFIQ2_KL27.27 GBDownload
IQ2_KL/Kimi-K2.7-Code-IQ2_KL-00006-of-00014.ggufGGUFIQ2_KL27.27 GBDownload
IQ2_KL/Kimi-K2.7-Code-IQ2_KL-00007-of-00014.ggufGGUFIQ2_KL27.27 GBDownload
IQ2_KL/Kimi-K2.7-Code-IQ2_KL-00008-of-00014.ggufGGUFIQ2_KL27.27 GBDownload
IQ2_KL/Kimi-K2.7-Code-IQ2_KL-00009-of-00014.ggufGGUFIQ2_KL27.27 GBDownload
IQ2_KL/Kimi-K2.7-Code-IQ2_KL-00010-of-00014.ggufGGUFIQ2_KL27.27 GBDownload
IQ2_KL/Kimi-K2.7-Code-IQ2_KL-00011-of-00014.ggufGGUFIQ2_KL27.27 GBDownload
IQ2_KL/Kimi-K2.7-Code-IQ2_KL-00012-of-00014.ggufGGUFIQ2_KL27.27 GBDownload
IQ2_KL/Kimi-K2.7-Code-IQ2_KL-00013-of-00014.ggufGGUFIQ2_KL27.27 GBDownload
IQ2_KL/Kimi-K2.7-Code-IQ2_KL-00014-of-00014.ggufGGUFIQ2_KL3.57 GBDownload
IQ2_KS/Kimi-K2.7-Code-IQ2_KS-00001-of-00014.ggufGGUFIQ2_KS6.6 MBDownload
IQ2_KS/Kimi-K2.7-Code-IQ2_KS-00002-of-00014.ggufGGUFIQ2_KS22.25 GBDownload
IQ2_KS/Kimi-K2.7-Code-IQ2_KS-00003-of-00014.ggufGGUFIQ2_KS22.34 GBDownload
IQ2_KS/Kimi-K2.7-Code-IQ2_KS-00004-of-00014.ggufGGUFIQ2_KS22.34 GBDownload
IQ2_KS/Kimi-K2.7-Code-IQ2_KS-00005-of-00014.ggufGGUFIQ2_KS22.34 GBDownload
IQ2_KS/Kimi-K2.7-Code-IQ2_KS-00006-of-00014.ggufGGUFIQ2_KS22.34 GBDownload
IQ2_KS/Kimi-K2.7-Code-IQ2_KS-00007-of-00014.ggufGGUFIQ2_KS22.34 GBDownload
IQ2_KS/Kimi-K2.7-Code-IQ2_KS-00008-of-00014.ggufGGUFIQ2_KS22.34 GBDownload
IQ2_KS/Kimi-K2.7-Code-IQ2_KS-00009-of-00014.ggufGGUFIQ2_KS22.34 GBDownload
IQ2_KS/Kimi-K2.7-Code-IQ2_KS-00010-of-00014.ggufGGUFIQ2_KS22.34 GBDownload
IQ2_KS/Kimi-K2.7-Code-IQ2_KS-00011-of-00014.ggufGGUFIQ2_KS22.34 GBDownload
IQ2_KS/Kimi-K2.7-Code-IQ2_KS-00012-of-00014.ggufGGUFIQ2_KS22.34 GBDownload
IQ2_KS/Kimi-K2.7-Code-IQ2_KS-00013-of-00014.ggufGGUFIQ2_KS22.34 GBDownload
IQ2_KS/Kimi-K2.7-Code-IQ2_KS-00014-of-00014.ggufGGUFIQ2_KS2.91 GBDownload
IQ2_KT/Kimi-K2.7-Code-IQ2_KT-00001-of-00014.ggufGGUFIQ2_KT6.6 MBDownload
IQ2_KT/Kimi-K2.7-Code-IQ2_KT-00002-of-00014.ggufGGUFIQ2_KT21.75 GBDownload
IQ2_KT/Kimi-K2.7-Code-IQ2_KT-00003-of-00014.ggufGGUFIQ2_KT21.77 GBDownload
IQ2_KT/Kimi-K2.7-Code-IQ2_KT-00004-of-00014.ggufGGUFIQ2_KT21.77 GBDownload
IQ2_KT/Kimi-K2.7-Code-IQ2_KT-00005-of-00014.ggufGGUFIQ2_KT21.77 GBDownload
IQ2_KT/Kimi-K2.7-Code-IQ2_KT-00006-of-00014.ggufGGUFIQ2_KT21.77 GBDownload
IQ2_KT/Kimi-K2.7-Code-IQ2_KT-00007-of-00014.ggufGGUFIQ2_KT21.77 GBDownload
IQ2_KT/Kimi-K2.7-Code-IQ2_KT-00008-of-00014.ggufGGUFIQ2_KT21.77 GBDownload
IQ2_KT/Kimi-K2.7-Code-IQ2_KT-00009-of-00014.ggufGGUFIQ2_KT21.77 GBDownload
IQ2_KT/Kimi-K2.7-Code-IQ2_KT-00010-of-00014.ggufGGUFIQ2_KT21.77 GBDownload
IQ2_KT/Kimi-K2.7-Code-IQ2_KT-00011-of-00014.ggufGGUFIQ2_KT21.77 GBDownload
IQ2_KT/Kimi-K2.7-Code-IQ2_KT-00012-of-00014.ggufGGUFIQ2_KT21.77 GBDownload
IQ2_KT/Kimi-K2.7-Code-IQ2_KT-00013-of-00014.ggufGGUFIQ2_KT21.77 GBDownload
IQ2_KT/Kimi-K2.7-Code-IQ2_KT-00014-of-00014.ggufGGUFIQ2_KT2.83 GBDownload
IQ3_KT/Kimi-K2.7-Code-IQ3_KT-00001-of-00014.ggufGGUFIQ3_KT6.6 MBDownload
IQ3_KT/Kimi-K2.7-Code-IQ3_KT-00002-of-00014.ggufGGUFIQ3_KT30.28 GBDownload
IQ3_KT/Kimi-K2.7-Code-IQ3_KT-00003-of-00014.ggufGGUFIQ3_KT31.61 GBDownload
IQ3_KT/Kimi-K2.7-Code-IQ3_KT-00004-of-00014.ggufGGUFIQ3_KT31.61 GBDownload
IQ3_KT/Kimi-K2.7-Code-IQ3_KT-00005-of-00014.ggufGGUFIQ3_KT31.61 GBDownload
IQ3_KT/Kimi-K2.7-Code-IQ3_KT-00006-of-00014.ggufGGUFIQ3_KT31.61 GBDownload
IQ3_KT/Kimi-K2.7-Code-IQ3_KT-00007-of-00014.ggufGGUFIQ3_KT31.61 GBDownload
IQ3_KT/Kimi-K2.7-Code-IQ3_KT-00008-of-00014.ggufGGUFIQ3_KT31.61 GBDownload
IQ3_KT/Kimi-K2.7-Code-IQ3_KT-00009-of-00014.ggufGGUFIQ3_KT31.61 GBDownload
IQ3_KT/Kimi-K2.7-Code-IQ3_KT-00010-of-00014.ggufGGUFIQ3_KT31.61 GBDownload
IQ3_KT/Kimi-K2.7-Code-IQ3_KT-00011-of-00014.ggufGGUFIQ3_KT31.61 GBDownload
IQ3_KT/Kimi-K2.7-Code-IQ3_KT-00012-of-00014.ggufGGUFIQ3_KT31.61 GBDownload
IQ3_KT/Kimi-K2.7-Code-IQ3_KT-00013-of-00014.ggufGGUFIQ3_KT31.61 GBDownload
IQ3_KT/Kimi-K2.7-Code-IQ3_KT-00014-of-00014.ggufGGUFIQ3_KT4.15 GBDownload

Model Details

Model IDgghfexp/Kimi-K2.7-Code-GGUF
Authorgghfexp
Pipelinetext-generation
Licenseother
Base modelmoonshotai/Kimi-K2.7-Code
Last modified2026-07-23T04:59:41.000Z

Model README

---

quantized_by: gghfez

pipeline_tag: text-generation

base_model:

  • moonshotai/Kimi-K2.7-Code

license: other

license_name: modified-mit

license_link: https://huggingface.co/moonshotai/Kimi-K2.7/blob/main/LICENSE

base_model_relation: quantized

tags:

  • mla
  • imatrix
  • conversational
  • ik_llama.cpp

---

imatrix Quantization of moonshotai/Kimi-K2.7

ik_llama.cpp quants of moonshotai/Kimi-K2.7 using Unsloth's imatrix and Ubergarm's quant recipes*.

*embedding and output tensors left at q8_0

The other quants in this collection REQUIRE ik_llama.cpp fork to support the ik's latest SOTA quants and optimizations! Do not download these big files and expect them to run on mainline vanilla llama.cpp, ollama, LM Studio, KoboldCpp, etc!

NOTE ik_llama.cpp can also run your existing GGUFs from AesSedai, unsloth, bartowski, mradermacher, etc

Some of ik's new quants are supported with Nexesenex/croco.cpp fork of KoboldCPP with Windows builds for CUDA 12.9. Also check for Windows builds by Thireus here. which have been CUDA 12.8.

These quants provide best in class perplexity for the given memory footprint.

The IQ2_KT is the most accurate 2-bit Kimi-K2.7-Code quant I've found on huggingface but it's slower to run.

Available quants

IQ2_KT - 264.5 GiB

Final estimate: PPL over 568 chunks for n_ctx=512 = 2.8960 +/- 0.01474 (+44.14% vs baseline)

IQ2_KS - 270.9 GiB

Final estimate: PPL over 568 chunks for n_ctx=512 = 2.9740 +/- 0.01518 (+48.02% vs baseline)

IQ2_KL - 329.7 GiB

Final estimate: PPL over 568 chunks for n_ctx=512 = 2.4417 +/- 0.01166 (+21.52% vs baseline)

IQ3_KT - 381.8 GiB

PPL Untested / don't have the hardware. Responds coherently to a few prompts.

References

ACK

Run gghfexp/Kimi-K2.7-Code-GGUF with guIDE

Download guIDE — the AI-native code editor with local LLM inference and 69 built-in tools.

Download guIDE → · Browse 524k+ models · Compare models

Source: Hugging Face · Compare models