Indexnusrefather/Erebus-RP-12B-Instruct-2608-v1-GGUF overview
Erebus RP 12B Instruct 2608 v1: Complex Finetune of Gemma 3 12b it Aimed at Enhancing Roleplay and Creative Writing. Erebus RP 12B Instruct 2608 v1 https://cdn…
Runs locally from ~5.60 GB disk (8 GB VRAM class GPUs with llama.cpp / guIDE).
Repository Files & Downloads
| File | Type | Quantization | Size | Link |
|---|---|---|---|---|
| Erebus-RP-12B-Instruct-2608-v1.BF16.gguf | GGUF | GGUF | 21.92 GB | Download |
| Erebus-RP-12B-Instruct-2608-v1.Q3_K_M.gguf | GGUF | GGUF | 5.60 GB | Download |
| Erebus-RP-12B-Instruct-2608-v1.Q4_K_M.gguf | GGUF | GGUF | 6.80 GB | Download |
| Erebus-RP-12B-Instruct-2608-v1.Q5_K_M.gguf | GGUF | GGUF | 7.87 GB | Download |
| Erebus-RP-12B-Instruct-2608-v1.Q6_K.gguf | GGUF | GGUF | 9.00 GB | Download |
| Erebus-RP-12B-Instruct-2608-v1.Q8_0.gguf | GGUF | GGUF | 11.65 GB | Download |
Model Details
| Model ID | Indexnusrefather/Erebus-RP-12B-Instruct-2608-v1-GGUF |
|---|---|
| Author | Indexnusrefather |
| Pipeline | text-generation |
| License | gemma |
| Base model | google/gemma-3-12b-it |
| Last modified | 2026-08-17T12:45:52.000Z |
Model README
---
license: gemma
language:
- en
base_model:
- google/gemma-3-12b-it
base_model_relation: finetune
pipeline_tag: text-generation
library_name: transformers
tags:
- GGUF
- RP
- Roleplay
- Creative
- Writer
- gemma3
- finetune
- ERP
- Instruct
- v1
- Creative writing
- experimental
- rich worldbuilding
---
Erebus-RP-12B-Instruct-2608-v1: Complex Finetune of Gemma 3 12b it Aimed at Enhancing Roleplay and Creative Writing.
-
!Erebus-RP-12B-Instruct-2608-v1
"Specialized dataset was used to aggressively make the model roleplay close to how bigger models do, resulting in longer messages and better track of the story."
---
Quick Overview:
Better roleplay than the base:
- Model was trained on carefully selected number of high quality chat logs, filtered only for long term conversations with proper assistant turns.
- Entire process was carefully controlled by me to ensure that model can change its writing style without overcooking, this version is a result of repeated attempts until I finally found the right setup.
- Reduced refusals due to dataset containing a number of explicit logs.
Quants(this time I will release safetensors and quants in different repos for more convenience):
- BF16: Overkill
- Q8_0: Highest quality, still overkill
- Q6_K: Extremely high quality, near lossless
- Q5_K_M: Very high quality, fast, recommended.
- Q4_K_M: High quality, very fast, saves a lot of space, recommended.
- Q3_K_M: Lower quality, fastest.
Note:
This model turned out pretty well, Nyx is more intelligent and better and instruction following, Erebus is more creative.
I didn't select gemma 4 12b as the base because it was a hell to work with, and was in my observations way more heavily RLed than the gemma 3 12b it, so I took the older generation as the base.
Next I'll probably work on finetuning Mellum 2 12B A2.5B Instruct, also might turn my attention back to Ministral 3 2512. either 3B or 8B I don't know yet,
Run Indexnusrefather/Erebus-RP-12B-Instruct-2608-v1-GGUF with guIDE
Download guIDE — the AI-native code editor with local LLM inference and 69 built-in tools.
Source: Hugging Face · Compare models