GraySoft
Projects Models Compare Cloud benchmarks FAQ Download guIDE →
Model Intelligence Sheet

ProCreations/grug-35b-mtp-gguf overview

grug 35b mtp gguf grug rock WITH prediction head inside. Q8 0 / Q5 K M / Q4 K M, mtp tensors included for llama.cpp speculative decoding needs build with qwen3…

ggufgrugmtpspeculative-decodingenbase_model:ProCreations/grug-35b-mtpbase_model:quantized:ProCreations/grug-35b-mtplicense:apache-2.0endpoints_compatibleregion:usconversational

Runs locally from ~857.6 MB disk (4 GB VRAM class GPUs with llama.cpp / guIDE).

Downloads
60
Likes
1
Pipeline

Repository Files & Downloads

4 GGUF files detected
Direct downloads for local inference
FileTypeQuantizationSizeLink
grug-35b-mtp-Q4_K_M.ggufGGUFQ4_K_M20.22 GBDownload
grug-35b-mtp-Q5_K_M.ggufGGUFQ5_K_M23.61 GBDownload
grug-35b-mtp-Q8_0.ggufGGUFQ8_035.21 GBDownload
mmproj-grug-35b-mtp-f16.ggufGGUFF16857.6 MBDownload

Model Details

Model IDProCreations/grug-35b-mtp-gguf
AuthorProCreations
Pipeline
Licenseapache-2.0
Base modelProCreations/grug-35b-mtp
Last modified2026-07-24T20:24:13.000Z

Model README

---

license: apache-2.0

base_model: ProCreations/grug-35b-mtp

tags: [grug, mtp, gguf, speculative-decoding]

language: [en]

---

grug-35b-mtp-gguf

grug rock WITH prediction head inside. Q8_0 / Q5_K_M / Q4_K_M, mtp tensors

included for llama.cpp speculative decoding (needs build with qwen3_5 MTP

support). draft head trained on grug data: t+2 agreement 68.1% (from-scratch head, experimental).

eye rock

caveman ask why grug blind. grug find eye.

download mmproj-grug-35b-mtp-f16.gguf beside any brain rock, then:

llama-server -m grug-35b-mtp-Q4_K_M.gguf \
  --mmproj mmproj-grug-35b-mtp-f16.gguf

eye rock same vision projector as v2.1 parent. MTP graft only touch prediction

head; vision tower unchanged. need recent llama.cpp with qwen3_5_moe + MTP

support.

main card: grug-35b-mtp.

no-MTP rocks: grug-35b-v2-gguf.

grug made by ProCreations.

Run ProCreations/grug-35b-mtp-gguf with guIDE

Download guIDE — the AI-native code editor with local LLM inference and 69 built-in tools.

Download guIDE → · Browse 524k+ models · Compare models

Source: Hugging Face · Compare models