Model Intelligence Sheet
toiar/Ri-Gemma-E2B-it-QAT-GGUF overview
toiar/Ri-Gemma-E2B-it-QAT-GGUF GGUF downloads for local AI inference with guIDE.
Runs locally from ~941.1 MB disk (4 GB VRAM class GPUs with llama.cpp / guIDE).
Repository Files & Downloads
5 GGUF files detected
Direct downloads for local inference
| File | Type | Quantization | Size | Link |
|---|---|---|---|---|
| ri-gemma-e2b-qat-part-2.BF16-mmproj.gguf | GGUF | GGUF | 941.1 MB | Download |
| ri-gemma-e2b-qat-part-2.F16.gguf | GGUF | GGUF | 8.67 GB | Download |
| ri-gemma-e2b-qat-part-2.Q4_K_M.gguf | GGUF | GGUF | 3.19 GB | Download |
| ri-gemma-e2b-qat-part-2.Q6_K.gguf | GGUF | GGUF | 3.58 GB | Download |
| ri-gemma-e2b-qat-part-2.Q8_0.gguf | GGUF | GGUF | 4.63 GB | Download |
Model Details
| Model ID | toiar/Ri-Gemma-E2B-it-QAT-GGUF |
|---|---|
| Author | toiar |
| Pipeline | text-generation |
| License | gemma |
| Base model | toiar/Ri-Gemma-E2B-IT-QAT-Khasi-Chatbot |
| Last modified | 2026-08-01T10:17:05.000Z |
Run toiar/Ri-Gemma-E2B-it-QAT-GGUF with guIDE
Download guIDE — the AI-native code editor with local LLM inference and 69 built-in tools.
Source: Hugging Face · Compare models