GraySoft
Projects Models Compare Cloud benchmarks FAQ Download guIDE →
Model Intelligence Sheet

ibm-granite/granite-vision-4.1-4b-GGUF overview

Granite Vision 4.1 4B GGUF NOTE This repository contains models that have been converted to the GGUF format with various quantizations from an IBM Granite base…

llama.cppgguflanguagegranite-4.1visionbase_model:ibm-granite/granite-vision-4.1-4bbase_model:quantized:ibm-granite/granite-vision-4.1-4blicense:apache-2.0endpoints_compatibleregion:usconversational

Runs locally from ~1.08 GB disk (4 GB VRAM class GPUs with llama.cpp / guIDE).

Downloads
17
Likes
1
Pipeline

Repository Files & Downloads

6 GGUF files detected
Direct downloads for local inference
FileTypeQuantizationSizeLink
granite-vision-4.1-4b-Q4_K_M.ggufGGUFQ4_K_M1.96 GBDownload
granite-vision-4.1-4b-Q5_K_M.ggufGGUFQ5_K_M2.27 GBDownload
granite-vision-4.1-4b-Q6_K.ggufGGUFQ6_K2.60 GBDownload
granite-vision-4.1-4b-Q8_0.ggufGGUFQ8_03.37 GBDownload
granite-vision-4.1-4b-bf16.ggufGGUFBF166.34 GBDownload
mmproj-model-f16.ggufGGUFF161.08 GBDownload

Model Details

Model IDibm-granite/granite-vision-4.1-4b-GGUF
Authoribm-granite
Pipeline
Licenseapache-2.0
Base modelibm-granite/granite-vision-4.1-4b
Last modified2026-06-24T17:40:40.000Z

Model README

---

license: apache-2.0

library_name: llama.cpp

tags:

  • language
  • granite-4.1
  • vision
  • gguf

base_model:

  • ibm-granite/granite-vision-4.1-4b

---

Granite-Vision-4.1-4B (GGUF)

> [!NOTE]

> This repository contains models that have been converted to the GGUF format with various quantizations from an IBM Granite base model.

>

> Please reference the base model's full model card here:

> https://huggingface.co/ibm-granite/granite-vision-4.1-4b

Requirements

  • llama.cpp build: b9534

Run ibm-granite/granite-vision-4.1-4b-GGUF with guIDE

Download guIDE — the AI-native code editor with local LLM inference and 69 built-in tools.

Download guIDE → · Browse 524k+ models · Compare models

Source: Hugging Face · Compare models