GraySoft
Projects Models Compare Cloud benchmarks FAQ Download guIDE →
Model Intelligence Sheet

Xlr8boi/Llama-3.2-1B-Instruct-GGUF overview

Llama 3.2 1B Instruct GGUF : GGUF This model was finetuned and converted to GGUF format using Unsloth https://github.com/unslothai/unsloth . Example usage : Fo…

ggufllamallama.cppunslothendpoints_compatibleregion:usconversational

Runs locally from ~770.3 MB disk (4 GB VRAM class GPUs with llama.cpp / guIDE).

Downloads
0
Likes
0
Pipeline
Author

Repository Files & Downloads

4 GGUF files detected
Direct downloads for local inference
FileTypeQuantizationSizeLink
Llama-3.2-1B.F16.ggufGGUFGGUF2.31 GBDownload
Llama-3.2-1B.Q4_K_M.ggufGGUFGGUF770.3 MBDownload
Llama-3.2-1B.Q5_K_M.ggufGGUFGGUF869.3 MBDownload
Llama-3.2-1B.Q8_0.ggufGGUFGGUF1.23 GBDownload

Model Details

Model IDXlr8boi/Llama-3.2-1B-Instruct-GGUF
AuthorXlr8boi
Pipeline
License
Base model
Last modified2026-07-05T14:17:30.000Z

Model README

---

tags:

  • gguf
  • llama.cpp
  • unsloth

---

Llama-3.2-1B-Instruct-GGUF : GGUF

This model was finetuned and converted to GGUF format using Unsloth.

Example usage:

  • For text only LLMs: llama-cli -hf Xlr8boi/Llama-3.2-1B-Instruct-GGUF --jinja
  • For multimodal models: llama-mtmd-cli -hf Xlr8boi/Llama-3.2-1B-Instruct-GGUF --jinja

Available Model files:

  • Llama-3.2-1B.Q5_K_M.gguf
  • Llama-3.2-1B.F16.gguf
  • Llama-3.2-1B.Q8_0.gguf
  • Llama-3.2-1B.Q4_K_M.gguf

This was trained 2x faster with Unsloth

<img src="https://raw.githubusercontent.com/unslothai/unsloth/main/images/unsloth%20made%20with%20love.png" width="200"/>

Run Xlr8boi/Llama-3.2-1B-Instruct-GGUF with guIDE

Download guIDE — the AI-native code editor with local LLM inference and 69 built-in tools.

Download guIDE → · Browse 524k+ models · Compare models

Source: Hugging Face · Compare models