GraySoft
Projects Models Compare Cloud benchmarks FAQ Download guIDE →
Model Intelligence Sheet

Adiuk/Qwen2.5-72B-Instruct-AdiTurbo-GGUF overview

Qwen2.5 72B Instruct AdiTurbo GGUF Low bit GGUF quantizations of Qwen/Qwen2.5 72B Instruct https://huggingface.co/Qwen/Qwen2.5 72B Instruct , produced with the…

ggufquantizedaditurbollama.cppbase_model:Qwen/Qwen2.5-72B-Instructbase_model:quantized:Qwen/Qwen2.5-72B-Instructlicense:otherendpoints_compatibleregion:usconversational

Runs locally from ~32.28 GB disk (32 GB+ VRAM class GPUs with llama.cpp / guIDE).

Downloads
0
Likes
0
Pipeline
Author

Repository Files & Downloads

2 GGUF files detected
Direct downloads for local inference
FileTypeQuantizationSizeLink
Qwen2.5-72B-Instruct-TQ3_0.ggufGGUFGGUF32.28 GBDownload
Qwen2.5-72B-Instruct-TQ4_0.ggufGGUFGGUF38.34 GBDownload

Model Details

Model IDAdiuk/Qwen2.5-72B-Instruct-AdiTurbo-GGUF
AuthorAdiuk
Pipeline
Licenseother
Base modelQwen/Qwen2.5-72B-Instruct
Last modified2026-06-21T21:15:24.000Z

Model README

---

license: other

base_model: Qwen/Qwen2.5-72B-Instruct

tags: [gguf, quantized, aditurbo, llama.cpp]

---

Qwen2.5-72B-Instruct-AdiTurbo-GGUF

Low-bit GGUF quantizations of Qwen/Qwen2.5-72B-Instruct,

produced with the AdiTurbo Engine (private llama.cpp fork, build 8661).

Quantized with a fresh 50-chunk wikitext-2 importance matrix. Intended for

offline/on-device inference (Eyla AIOS). See the AdiTurbo benchmark for

perplexity/throughput numbers.

Files

  • Qwen2.5-72B-Instruct-TQ3_0.gguf
  • Qwen2.5-72B-Instruct-TQ4_0.gguf

This is a derivative quantization; the original model's license (other) applies.

Run Adiuk/Qwen2.5-72B-Instruct-AdiTurbo-GGUF with guIDE

Download guIDE — the AI-native code editor with local LLM inference and 69 built-in tools.

Download guIDE → · Browse 524k+ models · Compare models

Source: Hugging Face · Compare models