Model Intelligence Sheet
Lufel6848/Gemma4-E4B-it-GGUF overview
Gemma4 E4B it GGUF : GGUF This model was finetuned and converted to GGUF format using Unsloth https://github.com/unslothai/unsloth . Example usage : For text o…
Runs locally from ~945.6 MB disk (4 GB VRAM class GPUs with llama.cpp / guIDE).
Repository Files & Downloads
6 GGUF files detected
Direct downloads for local inference
| File | Type | Quantization | Size | Link |
|---|---|---|---|---|
| gemma-4-E4B-it.BF16-mmproj.gguf | GGUF | GGUF | 945.6 MB | Download |
| gemma-4-E4B-it.BF16.gguf | GGUF | GGUF | 14.02 GB | Download |
| gemma-4-E4B-it.Q4_K_M.gguf | GGUF | GGUF | 4.97 GB | Download |
| gemma-4-E4B-it.Q5_K_M.gguf | GGUF | GGUF | 5.37 GB | Download |
| gemma-4-E4B-it.Q6_K.gguf | GGUF | GGUF | 5.79 GB | Download |
| gemma-4-E4B-it.Q8_0.gguf | GGUF | GGUF | 7.48 GB | Download |
Model Details
Model README
---
tags:
- gguf
- llama.cpp
- unsloth
- vision-language-model
---
Gemma4-E4B-it-GGUF : GGUF
This model was finetuned and converted to GGUF format using Unsloth.
Example usage:
- For text only LLMs:
llama-cli -hf Lufel6848/Gemma4-E4B-it-GGUF --jinja - For multimodal models:
llama-mtmd-cli -hf Lufel6848/Gemma4-E4B-it-GGUF --jinja
Available Model files:
gemma-4-E4B-it.Q8_0.ggufgemma-4-E4B-it.BF16.ggufgemma-4-E4B-it.Q6_K.ggufgemma-4-E4B-it.Q5_K_M.ggufgemma-4-E4B-it.Q4_K_M.ggufgemma-4-E4B-it.BF16-mmproj.gguf
⚠️ Ollama Note for Vision Models
Important: Ollama currently does not support separate mmproj files for vision models.
To create an Ollama model from this vision model:
- Place the
Modelfilein the same directory as the finetuned bf16 merged model - Run:
ollama create model_name -f ./Modelfile
(Replace model_name with your desired name)
This will create a unified bf16 model that Ollama can use.
This was trained 2x faster with Unsloth
Run Lufel6848/Gemma4-E4B-it-GGUF with guIDE
Download guIDE — the AI-native code editor with local LLM inference and 69 built-in tools.
Source: Hugging Face · Compare models