Model Intelligence Sheet
gguf-org/muse-glimmer-30b-gguf overview
muse glimmer 30b gguf execute tools and/or agentic workflow fit in 12GB vram/ram or less run it with ggk bash ggk server engine m muse glimmer 30b nvfp4.gguf m…
Runs locally from ~1.11 GB disk (4 GB VRAM class GPUs with llama.cpp / guIDE).
Repository Files & Downloads
Model Details
Model README
---
license: apache-2.0
base_model:
- meta-models/Muse-Glimmer-30B
---
muse-glimmer-30b-gguf
- execute tools and/or agentic workflow
- fit in 12GB vram/ram or less
run it with ggk
ggk server engine -- -m muse-glimmer-30b-nvfp4.gguf --mmproj mmproj-q4_0.gguf --jinja --spec-type draft-dflash -md dflash-q4_0.gguf --spec-draft-n-max 3 -ts 1,0
or run it with llama.cpp
./llama-server -m muse-glimmer-30b-nvfp4.gguf --mmproj mmproj-q4_0.gguf --jinja --spec-type draft-dflash -md dflash-q4_0.gguf --spec-draft-n-max 3 --host 127.0.0.1 --port 8888
or opt lmstudio, etc.
Run gguf-org/muse-glimmer-30b-gguf with guIDE
Download guIDE — the AI-native code editor with local LLM inference and 69 built-in tools.
Source: Hugging Face · Compare models