GraySoft
Projects Models Compare Cloud benchmarks FAQ Download guIDE →
Model Intelligence Sheet

Arki05/Qwen3.8-27B-GGUF-shards overview

Qwen3.8 27B — per type GGUF store qalloc store v2 This repo is a byte addressable quantization store , not a single quant. It holds Qwen3.8 27B quantized at ~4…

ggufquantizationbase_model:Qwen/Qwen3.8-27Bbase_model:quantized:Qwen/Qwen3.8-27Blicense:apache-2.0endpoints_compatibleregion:usconversational

Runs locally from ~13.0 MB disk (4 GB VRAM class GPUs with llama.cpp / guIDE).

Downloads
38,223
Likes
0
Pipeline
Author

Repository Files & Downloads

69 GGUF files detected
Direct downloads for local inference
FileTypeQuantizationSizeLink
calib/imatrix/qwen38_code.imatrix.ggufGGUFGGUF13.0 MBDownload
calib/imatrix/qwen38_encode_basis.imatrix.ggufGGUFGGUF13.0 MBDownload
calib/imatrix/qwen38_finemath.imatrix.ggufGGUFGGUF13.0 MBDownload
calib/imatrix/qwen38_fineweb-edu.imatrix.ggufGGUFGGUF13.0 MBDownload
calib/imatrix/qwen38_general-instruct.imatrix.ggufGGUFGGUF13.0 MBDownload
calib/imatrix/qwen38_nemotron-agentic.imatrix.ggufGGUFGGUF13.0 MBDownload
calib/imatrix/qwen38_reasoning.imatrix.ggufGGUFGGUF13.0 MBDownload
equivalents/UD-IQ1_M.ggufGGUFIQ1_M6.27 GBDownload
equivalents/UD-IQ1_S.ggufGGUFIQ1_S5.77 GBDownload
equivalents/UD-IQ2_S.ggufGGUFIQ2_S7.80 GBDownload
equivalents/UD-IQ2_XXS.ggufGGUFIQ2_XXS6.77 GBDownload
equivalents/UD-IQ3_S.ggufGGUFIQ3_S10.89 GBDownload
equivalents/UD-IQ3_XXS.ggufGGUFIQ3_XXS9.86 GBDownload
equivalents/UD-IQ4_XS.ggufGGUFIQ4_XS12.95 GBDownload
equivalents/UD-Q2_K_XL.ggufGGUFQ2_K_XL8.83 GBDownload
equivalents/UD-Q3_K_XL.ggufGGUFQ3_K_XL11.92 GBDownload
equivalents/UD-Q4_K_M.ggufGGUFQ4_K_M15.01 GBDownload
equivalents/UD-Q4_K_S.ggufGGUFQ4_K_S13.98 GBDownload
equivalents/UD-Q4_K_XL.ggufGGUFQ4_K_XL16.03 GBDownload
equivalents/UD-Q5_K_M.ggufGGUFQ5_K_M18.09 GBDownload
equivalents/UD-Q5_K_S.ggufGGUFQ5_K_S17.06 GBDownload
equivalents/UD-Q5_K_XL.ggufGGUFQ5_K_XL19.12 GBDownload
equivalents/UD-Q6_K.ggufGGUFQ6_K20.15 GBDownload
equivalents/UD-Q6_K_L.ggufGGUFQ6_K_L22.20 GBDownload
equivalents/UD-Q6_K_M.ggufGGUFQ6_K_M21.17 GBDownload
equivalents/UD-Q6_K_XL.ggufGGUFQ6_K_XL23.23 GBDownload
equivalents/UD-Q8_K_L.ggufGGUFQ8_K_L25.79 GBDownload
main.ggufGGUFGGUF20.6 MBDownload
types/IQ1_BN.ggufGGUFGGUF5.19 GBDownload
types/IQ1_KT.ggufGGUFGGUF7.92 GBDownload
types/IQ1_M.ggufGGUFGGUF7.91 GBDownload
types/IQ1_S.ggufGGUFGGUF7.38 GBDownload
types/IQ2_BN.ggufGGUFGGUF6.39 GBDownload
types/IQ2_K.ggufGGUFGGUF7.56 GBDownload
types/IQ2_KL.ggufGGUFGGUF8.57 GBDownload
types/IQ2_KS.ggufGGUFGGUF6.98 GBDownload
types/IQ2_KT.ggufGGUFGGUF8.99 GBDownload
types/IQ2_S.ggufGGUFGGUF10.21 GBDownload
types/IQ2_XS.ggufGGUFGGUF9.50 GBDownload
types/IQ2_XXS.ggufGGUFGGUF8.79 GBDownload
types/IQ3_K.ggufGGUFGGUF10.94 GBDownload
types/IQ3_KS.ggufGGUFGGUF10.16 GBDownload
types/IQ3_KT.ggufGGUFGGUF9.97 GBDownload
types/IQ3_S.ggufGGUFGGUF10.94 GBDownload
types/IQ3_XXS.ggufGGUFGGUF9.75 GBDownload
types/IQ4_K.ggufGGUFGGUF14.32 GBDownload
types/IQ4_KS.ggufGGUFGGUF13.54 GBDownload
types/IQ4_KSS.ggufGGUFGGUF12.75 GBDownload
types/IQ4_KT.ggufGGUFGGUF12.75 GBDownload
types/IQ4_NL.ggufGGUFGGUF14.32 GBDownload
types/IQ4_XS.ggufGGUFGGUF13.53 GBDownload
types/IQ5_K.ggufGGUFGGUF17.50 GBDownload
types/IQ5_KS.ggufGGUFGGUF16.72 GBDownload
types/IQ6_K.ggufGGUFGGUF21.08 GBDownload
types/MXFP4.ggufGGUFGGUF13.53 GBDownload
types/NVFP4.ggufGGUFGGUF14.32 GBDownload
types/Q2_K.ggufGGUFGGUF8.36 GBDownload
types/Q3_K.ggufGGUFGGUF10.94 GBDownload
types/Q4_0.ggufGGUFGGUF14.32 GBDownload
types/Q4_1.ggufGGUFGGUF15.91 GBDownload
types/Q4_K.ggufGGUFGGUF14.32 GBDownload
types/Q5_0.ggufGGUFGGUF17.50 GBDownload
types/Q5_1.ggufGGUFGGUF19.09 GBDownload
types/Q5_K.ggufGGUFGGUF17.50 GBDownload
types/Q6_0.ggufGGUFGGUF20.68 GBDownload
types/Q6_K.ggufGGUFGGUF20.88 GBDownload
types/Q8_0.ggufGGUFGGUF27.04 GBDownload
types/TQ1_0.ggufGGUFGGUF5.38 GBDownload
types/TQ2_0.ggufGGUFGGUF6.57 GBDownload

Model Details

Model IDArki05/Qwen3.8-27B-GGUF-shards
AuthorArki05
Pipeline
Licenseapache-2.0
Base modelQwen/Qwen3.8-27B
Last modified2026-08-27T04:03:08.000Z

Model README

---

license: apache-2.0

base_model: Qwen/Qwen3.8-27B

tags:

  • gguf
  • quantization

---

Qwen3.8-27B — per-type GGUF store (qalloc store-v2)

This repo is a byte-addressable quantization store, not a single quant.

It holds Qwen3.8-27B quantized at ~41 different types (mainline llama.cpp

and ik_llama.cpp families unified; per-type backend flag in the manifest):

  • types/<TYPE>.gguf — the whole model at one type. Every file is a normal,

directly runnable uniform quant (a mainline-typed file runs on stock

llama.cpp; ik types need ik_llama.cpp).

  • main.gguf — header KV metadata + all non-quantizable tensors (with the

upstream chat-template fix applied).

  • manifest.json (qalloc-manifest/v2) — the byte index: for every tensor

× type, the exact [offset, offset+bytes) range inside its per-type file.

A client can assemble any mixed-precision allocation with plain HTTP

Range requests — no server-side work.

  • calib/, eval/ — the calibration and evaluation corpus files used by

the damage-model program that drives allocation.

Assembly tooling and the allocation picker (exact MCKP solver over a

measured damage model) live in the qalloc project; the interactive picker

runs in the browser via WASM. The former v1 stores (per-tensor files, two

repos) are deprecated: this repo IS the store.

Run Arki05/Qwen3.8-27B-GGUF-shards with guIDE

Download guIDE — the AI-native code editor with local LLM inference and 69 built-in tools.

Download guIDE → · Browse 524k+ models · Compare models

Source: Hugging Face · Compare models