GraySoft
Projects Models Compare Cloud benchmarks FAQ Download guIDE →
Model Intelligence Sheet

batiai/Kimi-K2.7-Code-GGUF overview

Kimi K2.7 Code GGUF — Quantized by BatiAI <p align="center" <a href="https://flow.bati.ai" <img src="https://img.shields.io/badge/BatiFlow on device%20AI blue?…

llama.cppggufkimimoonshotquantizedbatiaimixture-of-expertscodingagenticfrontiertext-generationenkozhbase_model:moonshotai/Kimi-K2.7-Codebase_model:quantized:moonshotai/Kimi-K2.7-Codelicense:otherendpoints_compatibleregion:usimatrixconversational

Runs locally from ~2.01 GB disk (4 GB VRAM class GPUs with llama.cpp / guIDE).

Downloads
0
Likes
0
Pipeline
text-generation
Author

Repository Files & Downloads

23 GGUF files detected
Direct downloads for local inference
FileTypeQuantizationSizeLink
Kimi-K2.7-Code-IQ3_XXS-00001-of-00010.ggufGGUFIQ3_XXS40.03 GBDownload
Kimi-K2.7-Code-IQ3_XXS-00002-of-00010.ggufGGUFIQ3_XXS40.65 GBDownload
Kimi-K2.7-Code-IQ3_XXS-00003-of-00010.ggufGGUFIQ3_XXS40.61 GBDownload
Kimi-K2.7-Code-IQ3_XXS-00004-of-00010.ggufGGUFIQ3_XXS40.64 GBDownload
Kimi-K2.7-Code-IQ3_XXS-00005-of-00010.ggufGGUFIQ3_XXS40.65 GBDownload
Kimi-K2.7-Code-IQ3_XXS-00006-of-00010.ggufGGUFIQ3_XXS40.61 GBDownload
Kimi-K2.7-Code-IQ3_XXS-00007-of-00010.ggufGGUFIQ3_XXS40.64 GBDownload
Kimi-K2.7-Code-IQ3_XXS-00008-of-00010.ggufGGUFIQ3_XXS40.65 GBDownload
Kimi-K2.7-Code-IQ3_XXS-00009-of-00010.ggufGGUFIQ3_XXS40.61 GBDownload
Kimi-K2.7-Code-IQ3_XXS-00010-of-00010.ggufGGUFIQ3_XXS2.01 GBDownload
Kimi-K2.7-Code-IQ4_XS-00001-of-00013.ggufGGUFIQ4_XS41.18 GBDownload
Kimi-K2.7-Code-IQ4_XS-00002-of-00013.ggufGGUFIQ4_XS39.44 GBDownload
Kimi-K2.7-Code-IQ4_XS-00003-of-00013.ggufGGUFIQ4_XS39.45 GBDownload
Kimi-K2.7-Code-IQ4_XS-00004-of-00013.ggufGGUFIQ4_XS39.40 GBDownload
Kimi-K2.7-Code-IQ4_XS-00005-of-00013.ggufGGUFIQ4_XS39.44 GBDownload
Kimi-K2.7-Code-IQ4_XS-00006-of-00013.ggufGGUFIQ4_XS39.45 GBDownload
Kimi-K2.7-Code-IQ4_XS-00007-of-00013.ggufGGUFIQ4_XS39.40 GBDownload
Kimi-K2.7-Code-IQ4_XS-00008-of-00013.ggufGGUFIQ4_XS39.44 GBDownload
Kimi-K2.7-Code-IQ4_XS-00009-of-00013.ggufGGUFIQ4_XS39.45 GBDownload
Kimi-K2.7-Code-IQ4_XS-00010-of-00013.ggufGGUFIQ4_XS39.40 GBDownload
Kimi-K2.7-Code-IQ4_XS-00011-of-00013.ggufGGUFIQ4_XS39.44 GBDownload
Kimi-K2.7-Code-IQ4_XS-00012-of-00013.ggufGGUFIQ4_XS39.45 GBDownload
Kimi-K2.7-Code-IQ4_XS-00013-of-00013.ggufGGUFIQ4_XS33.75 GBDownload

Model Details

Model IDbatiai/Kimi-K2.7-Code-GGUF
Authorbatiai
Pipelinetext-generation
Licenseother
Base modelmoonshotai/Kimi-K2.7-Code
Last modified2026-08-02T05:24:14.000Z

Model README

---

language:

- en

- ko

- zh

license: other

license_name: modified-mit

license_link: https://huggingface.co/moonshotai/Kimi-K2.7-Code/blob/main/LICENSE

tags:

- gguf

- kimi

- moonshot

- quantized

- batiai

- mixture-of-experts

- coding

- agentic

- frontier

base_model: moonshotai/Kimi-K2.7-Code

pipeline_tag: text-generation

library_name: llama.cpp

---

Kimi-K2.7-Code GGUF — Quantized by BatiAI

<p align="center">

<a href="https://flow.bati.ai"><img src="https://img.shields.io/badge/BatiFlow-on--device%20AI-blue?style=for-the-badge&logo=apple" alt="BatiFlow"></a>

<a href="https://huggingface.co/moonshotai/Kimi-K2.7-Code"><img src="https://img.shields.io/badge/source-Moonshot%20official-orange?style=for-the-badge" alt="moonshot"></a>

<a href="#"><img src="https://img.shields.io/badge/1T--A32B-MoE-purple?style=for-the-badge" alt="MoE"></a>

</p>

> The coding upgrade to Kimi K2.6 — +21.8% on Kimi Code Bench v2, running on a 512GB Mac Studio.

> IQ3_XXS / IQ4_XS GGUF of moonshotai/Kimi-K2.7-Code (1T total / 32.6B active MoE, DeepSeek-V3-family architecture).

> Quantized directly from official Moonshot weights — code+multilingual imatrix, BatiAI-signed.

📦 Quantizations

| Quant | Size | Shards | Target |

|-------|------|--------|--------|

| IQ3_XXS | 394 GB (GiB: 367) | 10 | M3 Ultra 512GB Mac Studio |

| IQ4_XS | 546 GB (GiB: 509) | 13 | 512GB+ / multi-node / server |

Both built from official weights via a Q8_0 intermediate, quantized with a code + EN + KO + ZH imatrix (included: Kimi-K2.7-Code-imatrix.dat). Text-only (the vision tower of the K2.5-family checkpoint is not included; same as other K2 GGUFs).

✅ Verified (this build, IQ3_XXS) — captured greedy runs:

  • Math: 127+58185 (clean reasoning trace)
  • Korean: 서울 소개 + 김치·비빔밥·불고기 각 한 문장 — fluent, zero token-mixing or loops
  • Tool-call: {"tool":"get_weather","args":{"city":"부산"}} — exact JSON

🚀 Usage (llama.cpp — mainline, no fork needed)

hf download batiai/Kimi-K2.7-Code-GGUF "Kimi-K2.7-Code-IQ3_XXS-*.gguf" --local-dir ./k27

# llama.cpp auto-loads all shards from the first one
./llama-cli -m ./k27/Kimi-K2.7-Code-IQ3_XXS-00001-of-00010.gguf -ngl 99 -c 16384 \
  -p "Refactor this function and add tests."

Recommended sampling (Moonshot): --temp 1.0 --top-p 0.95 (thinking mode). Architecture is deepseek2 — supported by mainline llama.cpp out of the box. Ollama tags (batiai/kimi-k2.7-code) follow shortly.

📜 License

Modified MIT (Moonshot) — commercial use permitted; products exceeding 100M MAU / $20M monthly revenue must display "Kimi K2.7" attribution. Full text at the base model repo. Quantized weights redistributed under the same terms.

✨ What BatiAI did

  • Direct from official Moonshot weights (never a re-quant of third-party GGUFs)
  • Q8_0 intermediate + diverse imatrix (code/EN/KO/ZH) for balanced fidelity
  • Verified: load ✅ · math ✅ · Korean ✅ · tool-call JSON ✅ — BatiAI metadata-signed

BatiAI · on-device frontier AI · https://flow.bati.ai

Run batiai/Kimi-K2.7-Code-GGUF with guIDE

Download guIDE — the AI-native code editor with local LLM inference and 69 built-in tools.

Download guIDE → · Browse 524k+ models · Compare models

Source: Hugging Face · Compare models