GraySoft
Projects Models Compare Cloud benchmarks FAQ Download guIDE →
Model Intelligence Sheet

Adiuk/Llama-3.3-70B-Instruct-Q3_K_M-GGUF overview

Llama 3.3 70B Instruct Q3 K M GGUF Low bit GGUF quantizations of meta llama/Llama 3.3 70B Instruct https://huggingface.co/meta llama/Llama 3.3 70B Instruct , p…

ggufquantizedaditurbollama.cppbase_model:meta-llama/Llama-3.3-70B-Instructbase_model:quantized:meta-llama/Llama-3.3-70B-Instructlicense:llama3.3endpoints_compatibleregion:usimatrixconversational

Runs locally from ~31.91 GB disk (32 GB+ VRAM class GPUs with llama.cpp / guIDE).

Downloads
0
Likes
0
Pipeline
Author

Repository Files & Downloads

1 GGUF files detected
Direct downloads for local inference
FileTypeQuantizationSizeLink
Llama-3.3-70B-Instruct-GGUF-Q3_K_M.ggufGGUFQ3_K_M31.91 GBDownload

Model Details

Model IDAdiuk/Llama-3.3-70B-Instruct-Q3_K_M-GGUF
AuthorAdiuk
Pipeline
Licensellama3.3
Base modelmeta-llama/Llama-3.3-70B-Instruct
Last modified2026-06-21T21:21:19.000Z

Model README

---

license: llama3.3

base_model: meta-llama/Llama-3.3-70B-Instruct

tags: [gguf, quantized, aditurbo, llama.cpp]

---

Llama-3.3-70B-Instruct-Q3_K_M-GGUF

Low-bit GGUF quantizations of meta-llama/Llama-3.3-70B-Instruct,

produced with the AdiTurbo Engine (private llama.cpp fork, build 8661).

Quantized with a fresh 50-chunk wikitext-2 importance matrix. Intended for

offline/on-device inference (Eyla AIOS). See the AdiTurbo benchmark for

perplexity/throughput numbers.

Files

  • Llama-3.3-70B-Instruct-GGUF-Q3_K_M.gguf

This is a derivative quantization; the original model's license (llama3.3) applies.

Run Adiuk/Llama-3.3-70B-Instruct-Q3_K_M-GGUF with guIDE

Download guIDE — the AI-native code editor with local LLM inference and 69 built-in tools.

Download guIDE → · Browse 524k+ models · Compare models

Source: Hugging Face · Compare models