Model Intelligence Sheet
ibm-granite/granite-vision-4.1-4b-GGUF overview
Granite Vision 4.1 4B GGUF NOTE This repository contains models that have been converted to the GGUF format with various quantizations from an IBM Granite base…
Runs locally from ~1.08 GB disk (4 GB VRAM class GPUs with llama.cpp / guIDE).
Repository Files & Downloads
6 GGUF files detected
Direct downloads for local inference
| File | Type | Quantization | Size | Link |
|---|---|---|---|---|
| granite-vision-4.1-4b-Q4_K_M.gguf | GGUF | Q4_K_M | 1.96 GB | Download |
| granite-vision-4.1-4b-Q5_K_M.gguf | GGUF | Q5_K_M | 2.27 GB | Download |
| granite-vision-4.1-4b-Q6_K.gguf | GGUF | Q6_K | 2.60 GB | Download |
| granite-vision-4.1-4b-Q8_0.gguf | GGUF | Q8_0 | 3.37 GB | Download |
| granite-vision-4.1-4b-bf16.gguf | GGUF | BF16 | 6.34 GB | Download |
| mmproj-model-f16.gguf | GGUF | F16 | 1.08 GB | Download |
Model Details
| Model ID | ibm-granite/granite-vision-4.1-4b-GGUF |
|---|---|
| Author | ibm-granite |
| Pipeline | — |
| License | apache-2.0 |
| Base model | ibm-granite/granite-vision-4.1-4b |
| Last modified | 2026-06-24T17:40:40.000Z |
Model README
---
license: apache-2.0
library_name: llama.cpp
tags:
- language
- granite-4.1
- vision
- gguf
base_model:
- ibm-granite/granite-vision-4.1-4b
---
Granite-Vision-4.1-4B (GGUF)
> [!NOTE]
> This repository contains models that have been converted to the GGUF format with various quantizations from an IBM Granite base model.
>
> Please reference the base model's full model card here:
> https://huggingface.co/ibm-granite/granite-vision-4.1-4b
Requirements
- llama.cpp build: b9534
Run ibm-granite/granite-vision-4.1-4b-GGUF with guIDE
Download guIDE — the AI-native code editor with local LLM inference and 69 built-in tools.
Source: Hugging Face · Compare models