GraySoft
Projects Models Compare Cloud benchmarks FAQ Download guIDE →
Model Intelligence Sheet

liodon-ai/vicuna-13b-self_consistency_random_var_5-GGUF overview

vicuna 13b self consistency random var 5 — GGUF GGUF quantizations of Yuhan123/vicuna 13b self consistency random var 5 https://huggingface.co/Yuhan123/vicuna …

gguflocal-llmllama.cpplm-studioquantizedollamallamatext-generationbase_model:Yuhan123/vicuna-13b-self_consistency_random_var_5base_model:quantized:Yuhan123/vicuna-13b-self_consistency_random_var_5license:otherendpoints_compatibleregion:usconversational

Runs locally from ~4.52 GB disk (8 GB VRAM class GPUs with llama.cpp / guIDE).

Downloads
0
Likes
0
Pipeline
text-generation
Author

Repository Files & Downloads

7 GGUF files detected
Direct downloads for local inference
FileTypeQuantizationSizeLink
vicuna-13b-self_consistency_random_var_5-Q2_K.ggufGGUFQ2_K4.52 GBDownload
vicuna-13b-self_consistency_random_var_5-Q3_K_M.ggufGGUFQ3_K_M5.90 GBDownload
vicuna-13b-self_consistency_random_var_5-Q4_0.ggufGGUFQ4_06.86 GBDownload
vicuna-13b-self_consistency_random_var_5-Q4_K_M.ggufGGUFQ4_K_M7.33 GBDownload
vicuna-13b-self_consistency_random_var_5-Q5_K_M.ggufGGUFQ5_K_M8.60 GBDownload
vicuna-13b-self_consistency_random_var_5-Q6_K.ggufGGUFQ6_K9.95 GBDownload
vicuna-13b-self_consistency_random_var_5-Q8_0.ggufGGUFQ8_012.88 GBDownload

Model Details

Model IDliodon-ai/vicuna-13b-self_consistency_random_var_5-GGUF
Authorliodon-ai
Pipelinetext-generation
Licenseother
Base modelYuhan123/vicuna-13b-self_consistency_random_var_5
Last modified2026-10-06T14:03:00.000Z

Model README

---

license: other

base_model: Yuhan123/vicuna-13b-self_consistency_random_var_5

base_model_relation: quantized

library_name: gguf

pipeline_tag: text-generation

quantized_by: liodon-ai

tags:

  • gguf
  • local-llm
  • llama.cpp
  • lm-studio
  • quantized
  • ollama
  • llama

---

vicuna-13b-self_consistency_random_var_5 — GGUF

GGUF quantizations of Yuhan123/vicuna-13b-self_consistency_random_var_5, published by Liodon AI.

Quick Start

llama.cpp

llama-cli -hf liodon-ai/vicuna-13b-self_consistency_random_var_5-GGUF:Q4_K_M

Ollama

ollama run hf.co/liodon-ai/vicuna-13b-self_consistency_random_var_5-GGUF:Q4_K_M

LM Studio / Jan — search liodon-ai/vicuna-13b-self_consistency_random_var_5-GGUF and pick your quant.

Quants

| Quant | Size | VRAM est. | Notes |

|-------|------|-----------|-------|

| Q2_K | 4.85 GB | ~6 GB | 2-bit, smallest standard |

| Q3_K_M | 6.34 GB | ~7 GB | 3-bit, good for 8 GB VRAM |

| Q4_0 | 7.37 GB | ~8 GB | 4-bit original |

| Q4_K_M | 7.87 GB | ~9 GB | 4-bit (recommended sweet spot) |

| Q5_K_M | 9.23 GB | ~11 GB | 5-bit, high quality |

| Q6_K | 10.68 GB | ~12 GB | 6-bit, near-lossless |

| Q8_0 | 13.83 GB | ~16 GB | 8-bit, essentially lossless |

> For higher-quality sub-4-bit quants with iMatrix calibration, see liodon-ai/vicuna-13b-self_consistency_random_var_5-imatrix-GGUF

Source

Citation

@misc{liodonai_vicuna_13b_self_consistency_random_var_5_gguf,
  title        = {vicuna-13b-self_consistency_random_var_5 — GGUF},
  author       = {{Liodon AI}},
  year         = {2026},
  howpublished = {\url{https://huggingface.co/liodon-ai/vicuna-13b-self_consistency_random_var_5-GGUF}},
  note         = {GGUF quantization of Yuhan123/vicuna-13b-self_consistency_random_var_5}
}

---

Quantized by Liodon AI

Run liodon-ai/vicuna-13b-self_consistency_random_var_5-GGUF with guIDE

Download guIDE — the AI-native code editor with local LLM inference and 69 built-in tools.

Download guIDE → · Browse 524k+ models · Compare models

Source: Hugging Face · Compare models