Model Intelligence Sheet
LeadingPoint/R3-rerank-0.6b-GGUF overview
language: multilingual library name: llama.cpp base model: tencent/R3 rerank 0.6b tags: gguf llama.cpp embedding quantization R3 rerank 0.6b GGUF GGUF quantiza…
Runs locally from ~378.1 MB disk (4 GB VRAM class GPUs with llama.cpp / guIDE).
Repository Files & Downloads
Model Details
| Model ID | LeadingPoint/R3-rerank-0.6b-GGUF |
|---|---|
| Author | LeadingPoint |
| Pipeline | — |
| License | — |
| Base model | tencent/R3-rerank-0.6b |
| Last modified | 2026-08-02T17:49:22.000Z |
Model README
---
language:
- multilingual
library_name: llama.cpp
base_model:
- tencent/R3-rerank-0.6b
tags:
- gguf
- llama.cpp
- embedding
- quantization
---
R3-rerank-0.6b GGUF
GGUF quantizations of tencent/R3-rerank-0.6b.
Original Model
https://huggingface.co/tencent/R3-rerank-0.6b
All credit for the original authors.
Files
| File | Description |
|------|-------------|
| R3-rerank-0.6b-Q4_K_M.gguf | Q4_K_M quantization |
Conversion
Converted using the latest llama.cpp tools.
Run LeadingPoint/R3-rerank-0.6b-GGUF with guIDE
Download guIDE — the AI-native code editor with local LLM inference and 69 built-in tools.
Source: Hugging Face · Compare models