GraySoft
Projects Models Compare Cloud benchmarks FAQ Download guIDE →
Model Intelligence Sheet

Indexnusrefather/Erebus-RP-12B-Instruct-2608-v1-GGUF overview

Erebus RP 12B Instruct 2608 v1: Complex Finetune of Gemma 3 12b it Aimed at Enhancing Roleplay and Creative Writing. Erebus RP 12B Instruct 2608 v1 https://cdn…

transformersggufGGUFRPRoleplayCreativeWritergemma3finetuneERPInstructv1Creative writingexperimentalrich worldbuildingtext-generationenbase_model:google/gemma-3-12b-itbase_model:finetune:google/gemma-3-12b-itlicense:gemmaendpoints_compatibleregion:usconversational

Runs locally from ~5.60 GB disk (8 GB VRAM class GPUs with llama.cpp / guIDE).

Downloads
352
Likes
0
Pipeline
text-generation

Repository Files & Downloads

6 GGUF files detected
Direct downloads for local inference
FileTypeQuantizationSizeLink
Erebus-RP-12B-Instruct-2608-v1.BF16.ggufGGUFGGUF21.92 GBDownload
Erebus-RP-12B-Instruct-2608-v1.Q3_K_M.ggufGGUFGGUF5.60 GBDownload
Erebus-RP-12B-Instruct-2608-v1.Q4_K_M.ggufGGUFGGUF6.80 GBDownload
Erebus-RP-12B-Instruct-2608-v1.Q5_K_M.ggufGGUFGGUF7.87 GBDownload
Erebus-RP-12B-Instruct-2608-v1.Q6_K.ggufGGUFGGUF9.00 GBDownload
Erebus-RP-12B-Instruct-2608-v1.Q8_0.ggufGGUFGGUF11.65 GBDownload

Model Details

Model IDIndexnusrefather/Erebus-RP-12B-Instruct-2608-v1-GGUF
AuthorIndexnusrefather
Pipelinetext-generation
Licensegemma
Base modelgoogle/gemma-3-12b-it
Last modified2026-08-17T12:45:52.000Z

Model README

---

license: gemma

language:

  • en

base_model:

  • google/gemma-3-12b-it

base_model_relation: finetune

pipeline_tag: text-generation

library_name: transformers

tags:

  • GGUF
  • RP
  • Roleplay
  • Creative
  • Writer
  • gemma3
  • finetune
  • ERP
  • Instruct
  • v1
  • Creative writing
  • experimental
  • rich worldbuilding

---

Erebus-RP-12B-Instruct-2608-v1: Complex Finetune of Gemma 3 12b it Aimed at Enhancing Roleplay and Creative Writing.

-

!Erebus-RP-12B-Instruct-2608-v1

"Specialized dataset was used to aggressively make the model roleplay close to how bigger models do, resulting in longer messages and better track of the story."

---

Quick Overview:

Better roleplay than the base:

  • Model was trained on carefully selected number of high quality chat logs, filtered only for long term conversations with proper assistant turns.
  • Entire process was carefully controlled by me to ensure that model can change its writing style without overcooking, this version is a result of repeated attempts until I finally found the right setup.
  • Reduced refusals due to dataset containing a number of explicit logs.

Quants(this time I will release safetensors and quants in different repos for more convenience):

  • BF16: Overkill
  • Q8_0: Highest quality, still overkill
  • Q6_K: Extremely high quality, near lossless
  • Q5_K_M: Very high quality, fast, recommended.
  • Q4_K_M: High quality, very fast, saves a lot of space, recommended.
  • Q3_K_M: Lower quality, fastest.

Note:

This model turned out pretty well, Nyx is more intelligent and better and instruction following, Erebus is more creative.

I didn't select gemma 4 12b as the base because it was a hell to work with, and was in my observations way more heavily RLed than the gemma 3 12b it, so I took the older generation as the base.

Next I'll probably work on finetuning Mellum 2 12B A2.5B Instruct, also might turn my attention back to Ministral 3 2512. either 3B or 8B I don't know yet,

Run Indexnusrefather/Erebus-RP-12B-Instruct-2608-v1-GGUF with guIDE

Download guIDE — the AI-native code editor with local LLM inference and 69 built-in tools.

Download guIDE → · Browse 524k+ models · Compare models

Source: Hugging Face · Compare models