Model Intelligence Sheet
ggml-org/GLM-4.5-Air-GGUF overview
GLM 4.5 Air Run with https://llama.app bash llama serve hf ggml org/GLM 4.5 Air GGUF Source models https://huggingface.co/zai org/GLM 4.5 Air TODOs add info IM…
Runs locally from ~8.9 MB disk (4 GB VRAM class GPUs with llama.cpp / guIDE).
Repository Files & Downloads
6 GGUF files detected
Direct downloads for local inference
| File | Type | Quantization | Size | Link |
|---|---|---|---|---|
| GLM-4.5-Air-Q4_K_M-00001-of-00002.gguf | GGUF | Q4_K_M | 8.9 MB | Download |
| GLM-4.5-Air-Q4_K_M-00002-of-00002.gguf | GGUF | Q4_K_M | 59.25 GB | Download |
| GLM-4.5-Air-Q8_0-00001-of-00002.gguf | GGUF | Q8_0 | 8.9 MB | Download |
| GLM-4.5-Air-Q8_0-00002-of-00002.gguf | GGUF | Q8_0 | 105.80 GB | Download |
| mtp-GLM-4.5-Air-Q4_0.gguf | GGUF | Q4_0 | 2.56 GB | Download |
| mtp-GLM-4.5-Air-Q8_0.gguf | GGUF | Q8_0 | 4.82 GB | Download |
Model Details
| Model ID | ggml-org/GLM-4.5-Air-GGUF |
|---|---|
| Author | ggml-org |
| Pipeline | text-generation |
| License | mit |
| Base model | zai-org/GLM-4.5-Air |
| Last modified | 2026-08-25T11:02:51.000Z |
Model README
---
license: mit
pipeline_tag: text-generation
tags:
- gguf
- quantized
base_model:
- zai-org/GLM-4.5-Air
---
GLM-4.5-Air
Run with https://llama.app
llama serve -hf ggml-org/GLM-4.5-Air-GGUF
Source models
- https://huggingface.co/zai-org/GLM-4.5-Air
TODOs
- add info
> [!IMPORTANT]
> This model is automatically converted using https://github.com/ggml-org/convert
Run ggml-org/GLM-4.5-Air-GGUF with guIDE
Download guIDE — the AI-native code editor with local LLM inference and 69 built-in tools.
Source: Hugging Face · Compare models