GraySoft
Projects Models Compare Cloud benchmarks FAQ Download guIDE →
Model Intelligence Sheet

KeinNiemand/Kuwutu-7B-CYOA-GGUF-v2 overview

Kuwutu 7B CYOA GGUF v2 Mainline llama.cpp compatible GGUF quantizations of Kuwutu 7B CYOA v2 https://huggingface.co/KeinNiemand/Kuwutu 7B CYOA v2?not for all a…

ggufquantizationqwen2chatmlinteractive-fictioncyoansfwexplicitnot-for-all-audiencestext-generationenbase_model:KeinNiemand/Kuwutu-7B-CYOA-v2base_model:quantized:KeinNiemand/Kuwutu-7B-CYOA-v2license:cc-by-nc-4.0endpoints_compatibleregion:usconversational

Runs locally from ~2.67 GB disk (4 GB VRAM class GPUs with llama.cpp / guIDE).

Downloads
0
Likes
0
Pipeline
text-generation

Repository Files & Downloads

9 GGUF files detected
Direct downloads for local inference
FileTypeQuantizationSizeLink
Kuwutu-7B-CYOA-v2-IQ2_M.ggufGGUFIQ2_M2.67 GBDownload
Kuwutu-7B-CYOA-v2-IQ3_M.ggufGGUFIQ3_M3.42 GBDownload
Kuwutu-7B-CYOA-v2-IQ4_XS.ggufGGUFIQ4_XS3.97 GBDownload
Kuwutu-7B-CYOA-v2-Q3_K_M.ggufGGUFQ3_K_M3.59 GBDownload
Kuwutu-7B-CYOA-v2-Q4_K_M.ggufGGUFQ4_K_M4.37 GBDownload
Kuwutu-7B-CYOA-v2-Q5_K_M.ggufGGUFQ5_K_M5.08 GBDownload
Kuwutu-7B-CYOA-v2-Q6_K.ggufGGUFQ6_K5.83 GBDownload
Kuwutu-7B-CYOA-v2-Q8_0.ggufGGUFQ8_07.55 GBDownload
Kuwutu-7B-CYOA-v2-bf16.ggufGGUFBF1614.20 GBDownload

Model Details

Model IDKeinNiemand/Kuwutu-7B-CYOA-GGUF-v2
AuthorKeinNiemand
Pipelinetext-generation
Licensecc-by-nc-4.0
Base modelKeinNiemand/Kuwutu-7B-CYOA-v2
Last modified2026-07-15T08:39:51.000Z

Model README

---

base_model: KeinNiemand/Kuwutu-7B-CYOA-v2

base_model_relation: quantized

license: cc-by-nc-4.0

pipeline_tag: text-generation

language:

- en

tags:

- gguf

- quantization

- qwen2

- chatml

- interactive-fiction

- cyoa

- nsfw

- explicit

- not-for-all-audiences

---

Kuwutu-7B-CYOA-GGUF-v2

Mainline llama.cpp-compatible GGUF quantizations of Kuwutu-7B-CYOA-v2, a model for CYOA-style interactive fiction.

See the full model card for model details.

Files

  • Kuwutu-7B-CYOA-v2-bf16.gguf
  • Kuwutu-7B-CYOA-v2-Q8_0.gguf
  • Kuwutu-7B-CYOA-v2-Q6_K.gguf
  • Kuwutu-7B-CYOA-v2-Q5_K_M.gguf
  • Kuwutu-7B-CYOA-v2-Q4_K_M.gguf
  • Kuwutu-7B-CYOA-v2-Q3_K_M.gguf
  • Kuwutu-7B-CYOA-v2-IQ4_XS.gguf
  • Kuwutu-7B-CYOA-v2-IQ3_M.gguf
  • Kuwutu-7B-CYOA-v2-IQ2_M.gguf

The quantizations use an importance matrix generated from representative project calibration text.

Run KeinNiemand/Kuwutu-7B-CYOA-GGUF-v2 with guIDE

Download guIDE — the AI-native code editor with local LLM inference and 69 built-in tools.

Download guIDE → · Browse 524k+ models · Compare models

Source: Hugging Face · Compare models