GraySoft
Projects Models Compare Cloud benchmarks FAQ Download guIDE →
Model Intelligence Sheet

prithivMLmods/oMEGA-4B-SpatialThink-0804-GGUF overview

oMEGA 4B SpatialThink 0804 GGUF oMEGA 4B SpatialThink 0804 is a vision language model built on top of Qwen/Qwen3 VL 4B Instruct and fine tuned for spatial reas…

transformersgguftext-generation-inferencespatial-reasoningllama-cppvision-languagemultimodalimage-captioningvisual-question-answeringconditional-generationvisionlanguage-modelsftfine-grained-captioningcomputer-visionimage-text-to-textendataset:prithivMLmods/OpenCaption-FineGraineddataset:remyxai/SpaceThinkerbase_model:prithivMLmods/oMEGA-4B-SpatialThink-0804base_model:quantized:prithivMLmods/oMEGA-4B-SpatialThink-0804license:apache-2.0endpoints_compatibleregion:us

Runs locally from ~432.9 MB disk (4 GB VRAM class GPUs with llama.cpp / guIDE).

Downloads
0
Likes
2
Pipeline
image-text-to-text

Repository Files & Downloads

14 GGUF files detected
Direct downloads for local inference
FileTypeQuantizationSizeLink
oMEGA-4B-SpatialThink-0804.BF16.ggufGGUFGGUF7.50 GBDownload
oMEGA-4B-SpatialThink-0804.F16.ggufGGUFGGUF7.50 GBDownload
oMEGA-4B-SpatialThink-0804.Q3_K_L.ggufGGUFGGUF2.09 GBDownload
oMEGA-4B-SpatialThink-0804.Q3_K_M.ggufGGUFGGUF1.93 GBDownload
oMEGA-4B-SpatialThink-0804.Q3_K_S.ggufGGUFGGUF1.76 GBDownload
oMEGA-4B-SpatialThink-0804.Q4_K_M.ggufGGUFGGUF2.33 GBDownload
oMEGA-4B-SpatialThink-0804.Q4_K_S.ggufGGUFGGUF2.22 GBDownload
oMEGA-4B-SpatialThink-0804.Q5_K_M.ggufGGUFGGUF2.69 GBDownload
oMEGA-4B-SpatialThink-0804.Q5_K_S.ggufGGUFGGUF2.63 GBDownload
oMEGA-4B-SpatialThink-0804.Q6_K.ggufGGUFGGUF3.08 GBDownload
oMEGA-4B-SpatialThink-0804.Q8_0.ggufGGUFGGUF3.99 GBDownload
oMEGA-4B-SpatialThink-0804.mmproj-bf16.ggufGGUFBF16800.4 MBDownload
oMEGA-4B-SpatialThink-0804.mmproj-f16.ggufGGUFF16800.4 MBDownload
oMEGA-4B-SpatialThink-0804.mmproj-q8_0.ggufGGUFQ8_0432.9 MBDownload

Model Details

Model IDprithivMLmods/oMEGA-4B-SpatialThink-0804-GGUF
AuthorprithivMLmods
Pipelineimage-text-to-text
Licenseapache-2.0
Base modelprithivMLmods/oMEGA-4B-SpatialThink-0804
Last modified2026-08-04T09:35:44.000Z

Model README

---

base_model:

  • prithivMLmods/oMEGA-4B-SpatialThink-0804

tags:

  • text-generation-inference
  • spatial-reasoning
  • llama-cpp
  • vision-language
  • multimodal
  • image-captioning
  • visual-question-answering
  • conditional-generation
  • vision
  • language-model
  • sft
  • fine-grained-captioning
  • computer-vision

datasets:

  • prithivMLmods/OpenCaption-FineGrained
  • remyxai/SpaceThinker

license: apache-2.0

language:

  • en

pipeline_tag: image-text-to-text

library_name: transformers

---

oMEGA-4B-SpatialThink-0804-GGUF

> oMEGA-4B-SpatialThink-0804 is a vision-language model built on top of Qwen/Qwen3-VL-4B-Instruct and fine-tuned for spatial reasoning with concise notes for unfiltered vision tasks. The model is trained to produce concise yet informative reasoning for spatial understanding while maintaining strong image captioning capabilities. Training is based on remyxai's SpaceThinker and OpenCaption-FineGrained, enabling efficient spatial reasoning and detailed image understanding across diverse visual domains.

> [!NOTE]

> This model is an experimental release and may generate unexpected behaviors or reasoning artifacts in certain scenarios.

Model Files

File Name | Quant Type | File Size | File Link |

|-----------|------------|-----------|-----------|

| oMEGA-4B-SpatialThink-0804.BF16.gguf | BF16 | 8.05 GB | Download |

| oMEGA-4B-SpatialThink-0804.F16.gguf | F16 | 8.05 GB | Download |

| oMEGA-4B-SpatialThink-0804.Q3_K_L.gguf | Q3_K_L | 2.24 GB | Download |

| oMEGA-4B-SpatialThink-0804.Q3_K_M.gguf | Q3_K_M | 2.08 GB | Download |

| oMEGA-4B-SpatialThink-0804.Q3_K_S.gguf | Q3_K_S | 1.89 GB | Download |

| oMEGA-4B-SpatialThink-0804.Q4_K_M.gguf | Q4_K_M | 2.5 GB | Download |

| oMEGA-4B-SpatialThink-0804.Q4_K_S.gguf | Q4_K_S | 2.38 GB | Download |

| oMEGA-4B-SpatialThink-0804.Q5_K_M.gguf | Q5_K_M | 2.89 GB | Download |

| oMEGA-4B-SpatialThink-0804.Q5_K_S.gguf | Q5_K_S | 2.82 GB | Download |

| oMEGA-4B-SpatialThink-0804.Q6_K.gguf | Q6_K | 3.31 GB | Download |

| oMEGA-4B-SpatialThink-0804.Q8_0.gguf | Q8_0 | 4.28 GB | Download |

| oMEGA-4B-SpatialThink-0804.mmproj-bf16.gguf | mmproj-bf16 | 839 MB | Download |

| oMEGA-4B-SpatialThink-0804.mmproj-f16.gguf | mmproj-f16 | 839 MB | Download |

| oMEGA-4B-SpatialThink-0804.mmproj-q8_0.gguf | mmproj-q8_0 | 454 MB | Download |

llama.cpp

LLM inference in C/C++ — https://github.com/ggml-org/llama.cpp

Run prithivMLmods/oMEGA-4B-SpatialThink-0804-GGUF with guIDE

Download guIDE — the AI-native code editor with local LLM inference and 69 built-in tools.

Download guIDE → · Browse 524k+ models · Compare models

Source: Hugging Face · Compare models