GraySoft
Projects Models Compare Cloud benchmarks FAQ Download guIDE →
Model Intelligence Sheet

mrutkows/granite-4.1-3b-GGUF overview

granite 4.1 3b GGUF NOTE This repository contains models that have been converted to the GGUF format with various quantizations from an IBM Granite base model.…

llama.cppgguflanguagegranite-4.1base_model:ibm-granite/granite-4.1-3bbase_model:quantized:ibm-granite/granite-4.1-3blicense:apache-2.0endpoints_compatibleregion:usconversational

Runs locally from ~1.96 GB disk (4 GB VRAM class GPUs with llama.cpp / guIDE).

Downloads
1,028
Likes
0
Pipeline
Author

Repository Files & Downloads

5 GGUF files detected
Direct downloads for local inference
FileTypeQuantizationSizeLink
granite-4.1-3b-Q4_K_M.ggufGGUFQ4_K_M1.96 GBDownload
granite-4.1-3b-Q5_K_M.ggufGGUFQ5_K_M2.27 GBDownload
granite-4.1-3b-Q6_K.ggufGGUFQ6_K2.60 GBDownload
granite-4.1-3b-Q8_0.ggufGGUFQ8_03.37 GBDownload
granite-4.1-3b-bf16.ggufGGUFBF166.34 GBDownload

Model Details

Model IDmrutkows/granite-4.1-3b-GGUF
Authormrutkows
Pipeline
Licenseapache-2.0
Base modelibm-granite/granite-4.1-3b
Last modified2026-08-17T19:23:29.000Z

Model README

---

license: apache-2.0

library_name: llama.cpp

tags:

  • language
  • granite-4.1
  • gguf

base_model:

  • ibm-granite/granite-4.1-3b

---

granite-4.1-3b (GGUF)

> [!NOTE]

> This repository contains models that have been converted to the GGUF format with various quantizations from an IBM Granite base model.

>

> Please reference the base model's full model card here:

> https://huggingface.co/ibm-granite/granite-4.1-3b

<!-- #### Requirements

  • llama.cpp build: $b9768$ -->

Run mrutkows/granite-4.1-3b-GGUF with guIDE

Download guIDE — the AI-native code editor with local LLM inference and 69 built-in tools.

Download guIDE → · Browse 524k+ models · Compare models

Source: Hugging Face · Compare models