GraySoft
Projects Models Compare Cloud benchmarks FAQ Download guIDE →
Model Intelligence Sheet

prithivMLmods/dots.mocr-GGUF overview

dots.mocr GGUF dots.mocr is an advanced multimodal OCR model developed by rednote hilab https://huggingface.co/dots studio/dots.mocr as the successor to dots.o…

transformersgguftext-generation-inferencellama-cppimage-to-textocrdocument-parselayouttableformulacustom_codeimage-text-to-textenzhmultilingualbase_model:dots-studio/dots.mocrbase_model:quantized:dots-studio/dots.mocrlicense:mitendpoints_compatibleregion:usconversational

Runs locally from ~821.3 MB disk (4 GB VRAM class GPUs with llama.cpp / guIDE).

Downloads
0
Likes
1
Pipeline
image-text-to-text

Repository Files & Downloads

14 GGUF files detected
Direct downloads for local inference
FileTypeQuantizationSizeLink
dots.mocr.BF16.ggufGGUFGGUF3.32 GBDownload
dots.mocr.F16.ggufGGUFGGUF3.32 GBDownload
dots.mocr.Q3_K_L.ggufGGUFGGUF935.0 MBDownload
dots.mocr.Q3_K_M.ggufGGUFGGUF881.6 MBDownload
dots.mocr.Q3_K_S.ggufGGUFGGUF821.3 MBDownload
dots.mocr.Q4_K_M.ggufGGUFGGUF1.04 GBDownload
dots.mocr.Q4_K_S.ggufGGUFGGUF1021.9 MBDownload
dots.mocr.Q5_K_M.ggufGGUFGGUF1.20 GBDownload
dots.mocr.Q5_K_S.ggufGGUFGGUF1.17 GBDownload
dots.mocr.Q6_K.ggufGGUFGGUF1.36 GBDownload
dots.mocr.Q8_0.ggufGGUFGGUF1.76 GBDownload
dots.mocr.mmproj-bf16.ggufGGUFBF162.35 GBDownload
dots.mocr.mmproj-f16.ggufGGUFF162.35 GBDownload
dots.mocr.mmproj-q8_0.ggufGGUFQ8_01.25 GBDownload

Model Details

Model IDprithivMLmods/dots.mocr-GGUF
AuthorprithivMLmods
Pipelineimage-text-to-text
Licensemit
Base modeldots-studio/dots.mocr
Last modified2026-08-28T07:50:04.000Z

Model README

---

license: mit

language:

  • en
  • zh
  • multilingual

base_model:

  • dots-studio/dots.mocr

library_name: transformers

tags:

  • text-generation-inference
  • llama-cpp
  • image-to-text
  • ocr
  • document-parse
  • layout
  • table
  • formula
  • transformers
  • custom_code

pipeline_tag: image-text-to-text

---

dots.mocr-GGUF

> dots.mocr is an advanced multimodal OCR model developed by rednote-hilab as the successor to dots.ocr, built on a 3B-parameter vision-language model foundation that extends beyond standard document parsing to unify layout detection, content recognition, structured graphics parsing, grounding, semantic understanding, and interactive dialogue within a single framework. It achieves state-of-the-art performance among models of comparable size across multiple benchmarks, including OmniDocBench (v1.5), olmOCR-bench (83.9%), and XDocParse, surpassing competing specialized models like MonkeyOCR-pro-3B, GLM-OCR, PaddleOCR-VL-1.5, and HuanyuanOCR, while approaching the performance of much larger general VLMs like Gemini 3 Pro. A distinctive capability of dots.mocr is its ability to parse structured graphics — including charts, UI layouts, scientific figures, chemical formulas, and logos — directly into SVG code, with a companion model dots.mocr-svg specifically optimized for this image-to-SVG task, achieving scores of 0.902 on UniSVG, 0.905 on ChartMimic, and 0.901 on ChemDraw. The model supports multilingual document parsing across 100+ languages, handles diverse document types, outputs structured JSON with bounding boxes and layout categories, and supports inference via both HuggingFace Transformers and vLLM (officially integrated since vLLM v0.11.0), with additional capabilities including web parsing, scene text spotting, and general visual question answering.

Model Files

File Name | Quant Type | File Size | File Link |

|-----------|------------|-----------|-----------|

| dots.mocr.BF16.gguf | BF16 | 3.56 GB | Download |

| dots.mocr.F16.gguf | F16 | 3.56 GB | Download |

| dots.mocr.Q3_K_L.gguf | Q3_K_L | 980 MB | Download |

| dots.mocr.Q3_K_M.gguf | Q3_K_M | 924 MB | Download |

| dots.mocr.Q3_K_S.gguf | Q3_K_S | 861 MB | Download |

| dots.mocr.Q4_K_M.gguf | Q4_K_M | 1.12 GB | Download |

| dots.mocr.Q4_K_S.gguf | Q4_K_S | 1.07 GB | Download |

| dots.mocr.Q5_K_M.gguf | Q5_K_M | 1.29 GB | Download |

| dots.mocr.Q5_K_S.gguf | Q5_K_S | 1.26 GB | Download |

| dots.mocr.Q6_K.gguf | Q6_K | 1.46 GB | Download |

| dots.mocr.Q8_0.gguf | Q8_0 | 1.89 GB | Download |

| dots.mocr.mmproj-bf16.gguf | mmproj-bf16 | 2.53 GB | Download |

| dots.mocr.mmproj-f16.gguf | mmproj-f16 | 2.53 GB | Download |

| dots.mocr.mmproj-q8_0.gguf | mmproj-q8_0 | 1.34 GB | Download |

llama.cpp

LLM inference in C/C++ — https://github.com/ggml-org/llama.cpp

Run prithivMLmods/dots.mocr-GGUF with guIDE

Download guIDE — the AI-native code editor with local LLM inference and 69 built-in tools.

Download guIDE → · Browse 524k+ models · Compare models

Source: Hugging Face · Compare models