liodon-ai/vicuna-13b-self_consistency_random_var_5-GGUF overview
vicuna 13b self consistency random var 5 — GGUF GGUF quantizations of Yuhan123/vicuna 13b self consistency random var 5 https://huggingface.co/Yuhan123/vicuna …
Runs locally from ~4.52 GB disk (8 GB VRAM class GPUs with llama.cpp / guIDE).
Repository Files & Downloads
| File | Type | Quantization | Size | Link |
|---|---|---|---|---|
| vicuna-13b-self_consistency_random_var_5-Q2_K.gguf | GGUF | Q2_K | 4.52 GB | Download |
| vicuna-13b-self_consistency_random_var_5-Q3_K_M.gguf | GGUF | Q3_K_M | 5.90 GB | Download |
| vicuna-13b-self_consistency_random_var_5-Q4_0.gguf | GGUF | Q4_0 | 6.86 GB | Download |
| vicuna-13b-self_consistency_random_var_5-Q4_K_M.gguf | GGUF | Q4_K_M | 7.33 GB | Download |
| vicuna-13b-self_consistency_random_var_5-Q5_K_M.gguf | GGUF | Q5_K_M | 8.60 GB | Download |
| vicuna-13b-self_consistency_random_var_5-Q6_K.gguf | GGUF | Q6_K | 9.95 GB | Download |
| vicuna-13b-self_consistency_random_var_5-Q8_0.gguf | GGUF | Q8_0 | 12.88 GB | Download |
Model Details
| Model ID | liodon-ai/vicuna-13b-self_consistency_random_var_5-GGUF |
|---|---|
| Author | liodon-ai |
| Pipeline | text-generation |
| License | other |
| Base model | Yuhan123/vicuna-13b-self_consistency_random_var_5 |
| Last modified | 2026-10-06T14:03:00.000Z |
Model README
---
license: other
base_model: Yuhan123/vicuna-13b-self_consistency_random_var_5
base_model_relation: quantized
library_name: gguf
pipeline_tag: text-generation
quantized_by: liodon-ai
tags:
- gguf
- local-llm
- llama.cpp
- lm-studio
- quantized
- ollama
- llama
---
vicuna-13b-self_consistency_random_var_5 — GGUF
GGUF quantizations of Yuhan123/vicuna-13b-self_consistency_random_var_5, published by Liodon AI.
Quick Start
llama.cpp
llama-cli -hf liodon-ai/vicuna-13b-self_consistency_random_var_5-GGUF:Q4_K_M
Ollama
ollama run hf.co/liodon-ai/vicuna-13b-self_consistency_random_var_5-GGUF:Q4_K_M
LM Studio / Jan — search liodon-ai/vicuna-13b-self_consistency_random_var_5-GGUF and pick your quant.
Quants
| Quant | Size | VRAM est. | Notes |
|-------|------|-----------|-------|
| Q2_K | 4.85 GB | ~6 GB | 2-bit, smallest standard |
| Q3_K_M | 6.34 GB | ~7 GB | 3-bit, good for 8 GB VRAM |
| Q4_0 | 7.37 GB | ~8 GB | 4-bit original |
| Q4_K_M | 7.87 GB | ~9 GB | 4-bit (recommended sweet spot) |
| Q5_K_M | 9.23 GB | ~11 GB | 5-bit, high quality |
| Q6_K | 10.68 GB | ~12 GB | 6-bit, near-lossless |
| Q8_0 | 13.83 GB | ~16 GB | 8-bit, essentially lossless |
> For higher-quality sub-4-bit quants with iMatrix calibration, see liodon-ai/vicuna-13b-self_consistency_random_var_5-imatrix-GGUF
Source
- Model: Yuhan123/vicuna-13b-self_consistency_random_var_5
- License: other
Citation
@misc{liodonai_vicuna_13b_self_consistency_random_var_5_gguf,
title = {vicuna-13b-self_consistency_random_var_5 — GGUF},
author = {{Liodon AI}},
year = {2026},
howpublished = {\url{https://huggingface.co/liodon-ai/vicuna-13b-self_consistency_random_var_5-GGUF}},
note = {GGUF quantization of Yuhan123/vicuna-13b-self_consistency_random_var_5}
}
---
Quantized by Liodon AI
Run liodon-ai/vicuna-13b-self_consistency_random_var_5-GGUF with guIDE
Download guIDE — the AI-native code editor with local LLM inference and 69 built-in tools.
Source: Hugging Face · Compare models