GraySoft
Projects Models Compare Cloud benchmarks FAQ Download guIDE →
Model Intelligence Sheet

mrutkows/granite-4.1-8b-GGUF overview

granite 4.1 8b GGUF NOTE This repository contains models that have been converted to the GGUF format with various quantizations from an IBM Granite base model.…

llama.cppgguflanguagegranite-4.1base_model:ibm-granite/granite-4.1-8bbase_model:quantized:ibm-granite/granite-4.1-8blicense:apache-2.0endpoints_compatibleregion:usconversational

Runs locally from ~4.98 GB disk (8 GB VRAM class GPUs with llama.cpp / guIDE).

Downloads
25
Likes
0
Pipeline
Author

Repository Files & Downloads

4 GGUF files detected
Direct downloads for local inference
FileTypeQuantizationSizeLink
granite-4.1-8b-Q4_K_M.ggufGGUFQ4_K_M4.98 GBDownload
granite-4.1-8b-Q5_K_M.ggufGGUFQ5_K_M5.82 GBDownload
granite-4.1-8b-Q8_0.ggufGGUFQ8_08.70 GBDownload
granite-4.1-8b-bf16.ggufGGUFBF1616.38 GBDownload

Model Details

Model IDmrutkows/granite-4.1-8b-GGUF
Authormrutkows
Pipeline
Licenseapache-2.0
Base modelibm-granite/granite-4.1-8b
Last modified2026-08-14T15:49:45.000Z

Model README

---

license: apache-2.0

library_name: llama.cpp

tags:

  • language
  • granite-4.1
  • gguf

base_model:

  • ibm-granite/granite-4.1-8b

---

granite-4.1-8b (GGUF)

> [!NOTE]

> This repository contains models that have been converted to the GGUF format with various quantizations from an IBM Granite base model.

>

> Please reference the base model's full model card here:

> https://huggingface.co/ibm-granite/granite-4.1-8b

<!-- #### Requirements

  • llama.cpp build: $b9768$ -->

Run mrutkows/granite-4.1-8b-GGUF with guIDE

Download guIDE — the AI-native code editor with local LLM inference and 69 built-in tools.

Download guIDE → · Browse 524k+ models · Compare models

Source: Hugging Face · Compare models