GraySoft
Projects Models Compare Cloud benchmarks FAQ Download guIDE →
Model Intelligence Sheet

LeadingPoint/R3-rerank-0.6b-GGUF overview

language: multilingual library name: llama.cpp base model: tencent/R3 rerank 0.6b tags: gguf llama.cpp embedding quantization R3 rerank 0.6b GGUF GGUF quantiza…

llama.cppggufembeddingquantizationmultilingualbase_model:tencent/R3-rerank-0.6bbase_model:quantized:tencent/R3-rerank-0.6bendpoints_compatibleregion:usconversational

Runs locally from ~378.1 MB disk (4 GB VRAM class GPUs with llama.cpp / guIDE).

Downloads
113
Likes
1
Pipeline

Repository Files & Downloads

3 GGUF files detected
Direct downloads for local inference
FileTypeQuantizationSizeLink
R3-rerank-0.6b-BF16.ggufGGUFBF161.12 GBDownload
R3-rerank-0.6b-Q4_K_M.ggufGGUFQ4_K_M378.1 MBDownload
R3-rerank-0.6b-Q8_0.ggufGGUFQ8_0609.5 MBDownload

Model Details

Model IDLeadingPoint/R3-rerank-0.6b-GGUF
AuthorLeadingPoint
Pipeline
License
Base modeltencent/R3-rerank-0.6b
Last modified2026-08-02T17:49:22.000Z

Model README

---

language:

  • multilingual

library_name: llama.cpp

base_model:

  • tencent/R3-rerank-0.6b

tags:

  • gguf
  • llama.cpp
  • embedding
  • quantization

---

R3-rerank-0.6b GGUF

GGUF quantizations of tencent/R3-rerank-0.6b.

Original Model

https://huggingface.co/tencent/R3-rerank-0.6b

All credit for the original authors.

Files

| File | Description |

|------|-------------|

| R3-rerank-0.6b-Q4_K_M.gguf | Q4_K_M quantization |

Conversion

Converted using the latest llama.cpp tools.

Run LeadingPoint/R3-rerank-0.6b-GGUF with guIDE

Download guIDE — the AI-native code editor with local LLM inference and 69 built-in tools.

Download guIDE → · Browse 524k+ models · Compare models

Source: Hugging Face · Compare models