GraySoft
Projects Models Compare Cloud benchmarks FAQ Download guIDE →
Model Intelligence Sheet

ggml-org/gemma-4-31B-GGUF overview

gemma 4 31B Run with https://llama.app bash llama serve hf ggml org/gemma 4 31B GGUF Source models https://huggingface.co/google/gemma 4 31B TODOs add info IMP…

ggufquantizedimage-text-to-textbase_model:google/gemma-4-31Bbase_model:quantized:google/gemma-4-31Blicense:apache-2.0endpoints_compatibleregion:us

Runs locally from ~30.39 GB disk (32 GB+ VRAM class GPUs with llama.cpp / guIDE).

Downloads
0
Likes
1
Pipeline
image-text-to-text
Author

Repository Files & Downloads

2 GGUF files detected
Direct downloads for local inference
FileTypeQuantizationSizeLink
gemma-4-31B-BF16.ggufGGUFBF1657.20 GBDownload
gemma-4-31B-Q8_0.ggufGGUFQ8_030.39 GBDownload

Model Details

Model IDggml-org/gemma-4-31B-GGUF
Authorggml-org
Pipelineimage-text-to-text
Licenseapache-2.0
Base modelgoogle/gemma-4-31B
Last modified2026-07-16T18:26:09.000Z

Model README

---

license: apache-2.0

pipeline_tag: image-text-to-text

tags:

  • gguf
  • quantized

base_model:

  • google/gemma-4-31B

---

gemma-4-31B

Run with https://llama.app

llama serve -hf ggml-org/gemma-4-31B-GGUF

Source models

  • https://huggingface.co/google/gemma-4-31B

TODOs

  • add info

> [!IMPORTANT]

> This model is automatically converted using https://github.com/ggml-org/convert

Run ggml-org/gemma-4-31B-GGUF with guIDE

Download guIDE — the AI-native code editor with local LLM inference and 69 built-in tools.

Download guIDE → · Browse 524k+ models · Compare models

Source: Hugging Face · Compare models