GraySoft
Projects Models Compare Cloud benchmarks FAQ Download guIDE →
Model Intelligence Sheet

gguf-org/muse-glimmer-30b-gguf overview

muse glimmer 30b gguf execute tools and/or agentic workflow fit in 12GB vram/ram or less run it with ggk bash ggk server engine m muse glimmer 30b nvfp4.gguf m…

ggufbase_model:meta-models/Muse-Glimmer-30Bbase_model:quantized:meta-models/Muse-Glimmer-30Blicense:apache-2.0endpoints_compatibleregion:usconversational

Runs locally from ~1.11 GB disk (4 GB VRAM class GPUs with llama.cpp / guIDE).

Downloads
0
Likes
1
Pipeline
Author

Repository Files & Downloads

3 GGUF files detected
Direct downloads for local inference
FileTypeQuantizationSizeLink
dflash-q4_0.ggufGGUFQ4_01.35 GBDownload
mmproj-q4_0.ggufGGUFQ4_01.11 GBDownload
muse-glimmer-30b-nvfp4.ggufGGUFGGUF8.27 GBDownload

Model Details

Model IDgguf-org/muse-glimmer-30b-gguf
Authorgguf-org
Pipeline
Licenseapache-2.0
Base modelmeta-models/Muse-Glimmer-30B
Last modified2026-08-15T09:07:36.000Z

Model README

---

license: apache-2.0

base_model:

  • meta-models/Muse-Glimmer-30B

---

muse-glimmer-30b-gguf

  • execute tools and/or agentic workflow
  • fit in 12GB vram/ram or less

run it with ggk

ggk server engine -- -m muse-glimmer-30b-nvfp4.gguf --mmproj mmproj-q4_0.gguf --jinja --spec-type draft-dflash -md dflash-q4_0.gguf --spec-draft-n-max 3 -ts 1,0

or run it with llama.cpp

./llama-server -m muse-glimmer-30b-nvfp4.gguf --mmproj mmproj-q4_0.gguf --jinja --spec-type draft-dflash -md dflash-q4_0.gguf --spec-draft-n-max 3 --host 127.0.0.1 --port 8888

or opt lmstudio, etc.

Run gguf-org/muse-glimmer-30b-gguf with guIDE

Download guIDE — the AI-native code editor with local LLM inference and 69 built-in tools.

Download guIDE → · Browse 524k+ models · Compare models

Source: Hugging Face · Compare models