GraySoft
Projects Models About FAQ Contact Download guIDE →

ddh0/gemma-4-it-gguf 4.25bpw GGUF - Free GGUF Download is indexed on GraySoft with repository links, GGUF quant files, and Hugging Face metadata. This page helps you pick a local model for guIDE or other runtimes. See related models in the same shard below.

Model Intelligence Sheet

ddh0/gemma-4-it-gguf overview

These are miscellaneous GGUF quantizations of the instruct-tuned Gemma 4 series of models, released by Google. For more information about Gemma, you should refer to the original model cards. The chat template baked into these GGUFs is technically outdated, however, inference in llama.cpp should still work exactly as it should, thanks to these fixes: For the latest official chat template, refer to the original model repo.

ggufbase_model:google/gemma-4-26B-A4B-itbase_model:quantized:google/gemma-4-26B-A4B-itlicense:apache-2.0endpoints_compatibleregion:usimatrixconversational
ddh0/gemma-4-it-gguf visual
Downloads
98,039
Likes
1
Pipeline
Library
Visibility
Public
Access
Open

Repository Files & Downloads

26 files detected
Direct downloads for all repository files
FileTypeQuantizationSizeLink
gemma-4-26B-A4B-it-4.25bpw.gguf GGUF 12.50 GB Download
gemma-4-26B-A4B-it-4.74bpw.gguf GGUF 13.94 GB Download
gemma-4-26B-A4B-it-5.19bpw.gguf GGUF 15.27 GB Download
gemma-4-26B-A4B-it-6.70bpw.gguf GGUF 19.70 GB Download
gemma-4-26B-A4B-it-7.34bpw.gguf GGUF 21.58 GB Download
gemma-4-26B-A4B-it-8.51bpw.gguf GGUF 25.02 GB Download
gemma-4-31B-it-4.65bpw.gguf GGUF 16.63 GB Download
gemma-4-31B-it-5.02bpw.gguf GGUF 17.95 GB Download
gemma-4-31B-it-5.90bpw.gguf GGUF 21.11 GB Download
gemma-4-31B-it-6.02bpw.gguf GGUF 21.51 GB Download
gemma-4-31B-it-6.71bpw.gguf GGUF 23.98 GB Download
gemma-4-31B-it-7.63bpw.gguf GGUF 27.27 GB Download
gemma-4-31B-it-8.50bpw.gguf GGUF 30.39 GB Download
gemma-4-31B-it-bf16.gguf GGUF BF16 57.20 GB Download
gemma-4-E4B-it-6.99bpw.gguf GGUF 6.13 GB Download
gemma-4-E4B-it-7.15bpw.gguf GGUF 6.27 GB Download
gemma-4-E4B-it-7.60bpw.gguf GGUF 6.66 GB Download
gemma-4-E4B-it-8.19bpw.gguf GGUF 7.18 GB Download
gemma-4-E4B-it-8.76bpw.gguf GGUF 7.68 GB Download
gemma-4-E4B-it-f32.gguf GGUF F32 28.02 GB Download
mmproj-gemma-4-26B-A4B-it-f32.gguf GGUF F32 2.13 GB Download
mmproj-gemma-4-26B-A4B-it-q8_0.gguf GGUF 769.05 MB Download
mmproj-gemma-4-31B-it-f32.gguf GGUF F32 2.14 GB Download
mmproj-gemma-4-31B-it-q8_0.gguf GGUF 772.04 MB Download
mmproj-gemma-4-E4B-it-f32.gguf GGUF F32 1.78 GB Download
mmproj-gemma-4-E4B-it-q8_0.gguf GGUF 533.94 MB Download

Model Details Live

Model Slug
ddh0/gemma-4-it-gguf
Author
ddh0
Pipeline Task
Library
Created
2026-04-02
Last Modified
2026-04-15
Gated
No
Private
No
HF SHA
815b2a7ea89ce579ea9e93b13dc7bfcbf348c935
License
apache-2.0
Language
Unknown
Base Model
google/gemma-4-31B-it, google/gemma-4-26B-A4B-it, google/gemma-4-E4B-it, google/gemma-4-E2B-it

Metadata Inspector

Normalized metadata (stored in metadata_json)
{
  "metadata": {},
  "card_data": {
    "license": "apache-2.0",
    "base_model": [
      "google/gemma-4-31B-it",
      "google/gemma-4-26B-A4B-it",
      "google/gemma-4-E4B-it",
      "google/gemma-4-E2B-it"
    ],
    "frontmatter": {
      "license": "apache-2.0",
      "base_model": [
        "google/gemma-4-31B-it",
        "google/gemma-4-26B-A4B-it",
        "google/gemma-4-E4B-it",
        "google/gemma-4-E2B-it"
      ]
    },
    "hero_image_url": "",
    "summary": "These are miscellaneous GGUF quantizations of the instruct-tuned Gemma 4 series of models, released by Google. For more information about Gemma, you should refer to the original model cards. The chat template baked into these GGUFs is technically outdated, however, inference in llama.cpp should still work exactly as it should, thanks to these fixes: For the latest official chat template, refer to the original model repo.",
    "quick_links": [],
    "benchmark_table_html": "",
    "readme_markdown": "---\nlicense: apache-2.0\nbase_model:\n- google/gemma-4-31B-it\n- google/gemma-4-26B-A4B-it\n- google/gemma-4-E4B-it\n- google/gemma-4-E2B-it\n---\n\nThese are miscellaneous GGUF quantizations of the instruct-tuned Gemma 4 series of models, released by Google.\n\nFor more information about Gemma, you should refer to the original model cards.\n\nThe chat template baked into these GGUFs is technically outdated, however, inference in llama.cpp should still work exactly as it should, thanks to these fixes:\n- [llama.cpp#21704](https://github.com/ggml-org/llama.cpp/pull/21704): `common : better align to the updated official gemma4 template`\n- [llama.cpp#21760](https://github.com/ggml-org/llama.cpp/pull/21760): `common/gemma4 : handle parsing edge cases`\n\nFor the latest official chat template, refer to the original model repo.\n",
    "related_quantizations": []
  },
  "tags": [
    "gguf",
    "base_model:google/gemma-4-26B-A4B-it",
    "base_model:quantized:google/gemma-4-26B-A4B-it",
    "license:apache-2.0",
    "endpoints_compatible",
    "region:us",
    "imatrix",
    "conversational"
  ],
  "likes": 1,
  "downloads": 98039,
  "gated": false,
  "private": false,
  "last_modified": "2026-04-15T22:14:00.000Z",
  "created_at": "2026-04-02T21:34:12.000Z",
  "pipeline_tag": "",
  "library_name": ""
}
Source payload excerpt (from Hugging Face API)
{
  "_id": "69cee0d4711ef77392bdf3a4",
  "id": "ddh0/gemma-4-it-GGUF",
  "modelId": "ddh0/gemma-4-it-GGUF",
  "sha": "815b2a7ea89ce579ea9e93b13dc7bfcbf348c935",
  "createdAt": "2026-04-02T21:34:12.000Z",
  "lastModified": "2026-04-15T22:14:00.000Z",
  "author": "ddh0",
  "downloads": 98039,
  "likes": 1,
  "gated": false,
  "private": false,
  "pipeline_tag": "",
  "library_name": "",
  "siblings_count": 28
}