davidau/darkforest-20b-v3-ultra-quality-gguf Q6_K GGUF - Free GGUF Download is indexed on GraySoft with repository links, GGUF quant files, and Hugging Face metadata. This page helps you pick a local model for guIDE or other runtimes. See related models in the same shard below.
davidau/darkforest-20b-v3-ultra-quality-gguf overview
Ultra High Quality - 20 B Dark Forest Version 3.0 - 32 bit upscale Fully rebuilt from master files, including full merge(s) to maintain full 32 bit precision right up until it is compressed into GGUF files which results on a top to bottom upgrade. The result is superior performance in instruction following, reasoning, depth, nuance and emotion. NOTE: There are three original versions of "Dark Forest 20B", this is an upscale of the third version, with links below to 1st and 2nd versions also upscaled. On average this means a q4km operates at Q6 levels and Q6 and Q8 exceeds original model full precision performance. Perplexity drop (lower is better) is close to 10% (over 752 points for q4km) for all quants. That means precision has been enhanced for all 20 billion parameters which affects "brain density" / "function", instruction following and output quality. Imatrix quants to follow shortly. Settings: CHAT / ROLEPLAY and/or SMOOTHER operation of this model: In "KoboldCpp" or "oobabooga/text-generation-webui" or "Silly Tavern" ; Set the "Smoothingfactor" to 1.5 to 2.5 : in KoboldCpp -> Settings->Samplers->Advanced-> "SmoothF" : in text-generation-webui -> parameters -> lower right. : In Silly Tavern this is called: "Smoothing" NOTE: For "text-generation-webui" -> if using GGUFs you need to use "llamaHF" (which involves downloading some config files from the SOURCE version of this model) Source versions (and config files) of my models are here: https://huggingface.co/collections/DavidAU/d-au-source-files-for-gguf-exl2-awq-gptq-hqq-etc-etc-66b55cb8ba25f914cbf210be OTHER OPTIONS: For more details, including a list of enhancements see our other 32 bit upscale of "Space Whale 20B" rebuild here: [ https://huggingface.co/DavidAU/Psyonic-Cetacean-Ultra-Quality-20b-GGUF ] For Version 1 of Dark Forest Ultra Quality 32 bit upscale please go here: [ https://huggingface.co/DavidAU/Dark-Forest-V1-Ultra-Quality-20b-GGUF ] For Version 1 of Dark Forest Ultra Quality 32 bit upscale please go here: [ https://huggingface.co/TeeZee/DarkForest-20B-v2.0 ] Special thanks to "TEEZEE" for making a both fantasic models of "Dark Forest". Info from the original model card: Warning: This model can produce NSFW content! Results: For original model spec and information please visit: [ https://huggingface.co/TeeZee/DarkForest-20B-v3.0 ] Highest Quality Settings / Optimal Operation Guide / Parameters and Samplers This a "Class 2" model: For all settings used for this model (including specifics for its "class"), including example generation(s) and for advanced settings guide (which many times addresses any model issue(s)), including methods to improve model performance for all use case(s) as well as chat, roleplay and other use case(s) please see: [ https://huggingface.co/DavidAU/Maximizing-Model-Performance-All-Quants-Types-And-Full-Precision-by-SamplersParameters ] --- Special Thanks: --- Special thanks to all the following, and many more... All the model makers, fine tuners, mergers, and tweakers: Huggingface [ https://huggingface.co ] : LlamaCPP [ https://github.com/ggml-org/llama.cpp ] : Quant-Masters: Team Mradermacher, Bartowski, and many others: MergeKit [ https://github.com/arcee-ai/mergekit ] : Lmstudio [ https://lmstudio.ai/ ] : Text Generation Webui // KolboldCPP // SillyTavern:
Repository Files & Downloads
| File | Type | Quantization | Size | Link |
|---|---|---|---|---|
| DarkForest20B-V3-Ultra-Quality-IQ4_XS.gguf | GGUF | IQ4_XS | 10.01 GB | Download |
| DarkForest20B-V3-Ultra-Quality-Q2_k.gguf | GGUF | Q2_K | 6.91 GB | Download |
| DarkForest20B-V3-Ultra-Quality-Q3_k_l.gguf | GGUF | Q3_K_L | 9.90 GB | Download |
| DarkForest20B-V3-Ultra-Quality-Q3_k_m.gguf | GGUF | Q3_K_M | 9.04 GB | Download |
| DarkForest20B-V3-Ultra-Quality-Q3_k_s.gguf | GGUF | Q3_K_S | 8.06 GB | Download |
| DarkForest20B-V3-Ultra-Quality-Q4_k_m.gguf | GGUF | Q4_K_M | 11.22 GB | Download |
| DarkForest20B-V3-Ultra-Quality-Q4_k_s.gguf | GGUF | Q4_K_S | 10.59 GB | Download |
| DarkForest20B-V3-Ultra-Quality-Q5_k_s.gguf | GGUF | Q5_K_S | 12.83 GB | Download |
| DarkForest20B-V3-Ultra-Quality-Q6_k.gguf | GGUF | Q6_K | 15.28 GB | Download |
| DarkForest20B-V3-Ultra-Quality-Q8_0.gguf | GGUF | — | 19.79 GB | Download |
| DarkForest20B-V3-Ultra-Quality-q5_k_m.gguf | GGUF | Q5_K_M | 13.18 GB | Download |
Model Details Live
Metadata Inspector
Normalized metadata (stored in metadata_json)
{
"metadata": {},
"card_data": {
"license": "apache-2.0",
"language": [
"en"
],
"tags": [
"story",
"roleplay",
"creative",
"rp",
"fantasy",
"story telling",
"32 bit upscale",
"ultra high precision",
"nsfw",
"llama",
"llama-2"
],
"pipeline_tag": "text-generation",
"frontmatter": {
"license": "apache-2.0",
"language": [
"en"
],
"tags": [
"story",
"roleplay",
"creative",
"rp",
"fantasy",
"story telling",
"32 bit upscale",
"ultra high precision",
"nsfw",
"llama",
"llama-2"
],
"pipeline_tag": "text-generation"
},
"hero_image_url": "dark-forest.jpg",
"summary": "Ultra High Quality - 20 B Dark Forest Version 3.0 - 32 bit upscale Fully rebuilt from master files, including full merge(s) to maintain full 32 bit precision right up until it is compressed into GGUF files which results on a top to bottom upgrade. The result is superior performance in instruction following, reasoning, depth, nuance and emotion. NOTE: There are three original versions of \"Dark Forest 20B\", this is an upscale of the third version, with links below to 1st and 2nd versions also upscaled. On average this means a q4km operates at Q6 levels and Q6 and Q8 exceeds original model full precision performance. Perplexity drop (lower is better) is close to 10% (over 752 points for q4km) for all quants. That means precision has been enhanced for all 20 billion parameters which affects \"brain density\" / \"function\", instruction following and output quality. Imatrix quants to follow shortly. Settings: CHAT / ROLEPLAY and/or SMOOTHER operation of this model: In \"KoboldCpp\" or \"oobabooga/text-generation-webui\" or \"Silly Tavern\" ; Set the \"Smoothing_factor\" to 1.5 to 2.5 : in KoboldCpp -> Settings->Samplers->Advanced-> \"Smooth_F\" : in text-generation-webui -> parameters -> lower right. : In Silly Tavern this is called: \"Smoothing\" NOTE: For \"text-generation-webui\" -> if using GGUFs you need to use \"llama_HF\" (which involves downloading some config files from the SOURCE version of this model) Source versions (and config files) of my models are here: https://huggingface.co/collections/DavidAU/d-au-source-files-for-gguf-exl2-awq-gptq-hqq-etc-etc-66b55cb8ba25f914cbf210be OTHER OPTIONS: For more details, including a list of enhancements see our other 32 bit upscale of \"Space Whale 20B\" rebuild here: [ https://huggingface.co/DavidAU/Psyonic-Cetacean-Ultra-Quality-20b-GGUF ] For Version 1 of Dark Forest Ultra Quality 32 bit upscale please go here: [ https://huggingface.co/DavidAU/Dark-Forest-V1-Ultra-Quality-20b-GGUF ] For Version 1 of Dark Forest Ultra Quality 32 bit upscale please go here: [ https://huggingface.co/TeeZee/DarkForest-20B-v2.0 ] Special thanks to \"TEEZEE\" for making a both fantasic models of \"Dark Forest\". Info from the original model card: Warning: This model can produce NSFW content! Results: For original model spec and information please visit: [ https://huggingface.co/TeeZee/DarkForest-20B-v3.0 ] Highest Quality Settings / Optimal Operation Guide / Parameters and Samplers This a \"Class 2\" model: For all settings used for this model (including specifics for its \"class\"), including example generation(s) and for advanced settings guide (which many times addresses any model issue(s)), including methods to improve model performance for all use case(s) as well as chat, roleplay and other use case(s) please see: [ https://huggingface.co/DavidAU/Maximizing-Model-Performance-All-Quants-Types-And-Full-Precision-by-Samplers_Parameters ] --- Special Thanks: --- Special thanks to all the following, and many more... All the model makers, fine tuners, mergers, and tweakers: Huggingface [ https://huggingface.co ] : LlamaCPP [ https://github.com/ggml-org/llama.cpp ] : Quant-Masters: Team Mradermacher, Bartowski, and many others: MergeKit [ https://github.com/arcee-ai/mergekit ] : Lmstudio [ https://lmstudio.ai/ ] : Text Generation Webui // KolboldCPP // SillyTavern:",
"quick_links": [],
"benchmark_table_html": "",
"readme_markdown": "---\nlicense: apache-2.0\nlanguage:\n- en\ntags:\n- story\n- roleplay\n- creative\n- rp\n- fantasy\n- story telling\n- 32 bit upscale\n- ultra high precision\n- nsfw\n- llama\n- llama-2\npipeline_tag: text-generation\n---\n<B> Ultra High Quality - 20 B Dark Forest Version 3.0 - 32 bit upscale </b>\n\nFully rebuilt from master files, including full merge(s) to maintain full 32 bit precision right\nup until it is compressed into GGUF files which results on a top to bottom upgrade.\n\nThe result is superior performance in instruction following, reasoning, depth, nuance and emotion.\n\nNOTE: There are three original versions of \"Dark Forest 20B\", this is an upscale of the third version, with\nlinks below to 1st and 2nd versions also upscaled.\n\n<img src=\"dark-forest.jpg\">\n\nOn average this means a q4km operates at Q6 levels and Q6 and Q8 exceeds original model full precision performance.\n\nPerplexity drop (lower is better) is close to 10% (over 752 points for q4km) for all quants.\n\nThat means precision has been enhanced for all 20 billion parameters which affects \"brain density\" / \"function\", \ninstruction following and output quality.\n\nImatrix quants to follow shortly.\n\n<B>Settings: CHAT / ROLEPLAY and/or SMOOTHER operation of this model:</B>\n\nIn \"KoboldCpp\" or \"oobabooga/text-generation-webui\" or \"Silly Tavern\" ;\n\nSet the \"Smoothing_factor\" to 1.5 to 2.5 \n\n: in KoboldCpp -> Settings->Samplers->Advanced-> \"Smooth_F\"\n\n: in text-generation-webui -> parameters -> lower right.\n\n: In Silly Tavern this is called: \"Smoothing\"\n\n\nNOTE: For \"text-generation-webui\" \n\n-> if using GGUFs you need to use \"llama_HF\" (which involves downloading some config files from the SOURCE version of this model)\n\nSource versions (and config files) of my models are here:\n\nhttps://huggingface.co/collections/DavidAU/d-au-source-files-for-gguf-exl2-awq-gptq-hqq-etc-etc-66b55cb8ba25f914cbf210be\n\nOTHER OPTIONS:\n\n- Increase rep pen to 1.1 to 1.15 (you don't need to do this if you use \"smoothing_factor\")\n\n- If the interface/program you are using to run AI MODELS supports \"Quadratic Sampling\" (\"smoothing\") just make the adjustment as noted.\n\nFor more details, including a list of enhancements see our other 32 bit \nupscale of \"Space Whale 20B\" rebuild here:\n\n[ https://huggingface.co/DavidAU/Psyonic-Cetacean-Ultra-Quality-20b-GGUF ]\n\nFor Version 1 of Dark Forest Ultra Quality 32 bit upscale please go here:\n\n[ https://huggingface.co/DavidAU/Dark-Forest-V1-Ultra-Quality-20b-GGUF ]\n\nFor Version 1 of Dark Forest Ultra Quality 32 bit upscale please go here:\n\n[ https://huggingface.co/TeeZee/DarkForest-20B-v2.0 ]\n\nSpecial thanks to \"TEEZEE\" for making a both fantasic models of \"Dark Forest\".\n\n<b> Info from the original model card: </B>\n\nWarning: This model can produce NSFW content!\n\nResults:\n \n - main difference to v1.0 - model has much better sense of humor.\n - produces SFW nad NSFW content without issues, switches context seamlessly.\n - good at following instructions.\n - good at tracking multiple characters in one scene.\n - very creative, scenarios produced are mature and complicated, model doesn't shy from writing about PTSD, menatal issues or complicated relationships.\n - NSFW output is more creative and suprising than typical limaRP output.\n - definitely for mature audiences, not only because of vivid NSFW content but also because of overall maturity of stories it produces.\n - This is NOT Harry Potter level storytelling.\n\n\nFor original model spec and information please visit:\n\n[ https://huggingface.co/TeeZee/DarkForest-20B-v3.0 ]\n\n<B>Highest Quality Settings / Optimal Operation Guide / Parameters and Samplers</B>\n\nThis a \"Class 2\" model:\n\nFor all settings used for this model (including specifics for its \"class\"), including example generation(s) and for advanced settings guide (which many times addresses any model issue(s)), including methods to improve model performance for all use case(s) as well as chat, roleplay and other use case(s) please see:\n\n[ https://huggingface.co/DavidAU/Maximizing-Model-Performance-All-Quants-Types-And-Full-Precision-by-Samplers_Parameters ]\n\n---\n\n<h2>Special Thanks:</h2>\n\n---\n\nSpecial thanks to all the following, and many more...\n\nAll the model makers, fine tuners, mergers, and tweakers:\n- Provides the raw \"DNA\" for almost all my models.\n- Sources of model(s) can be found on the repo pages, especially the \"source\" repos with link(s) to the model creator(s).\n\nHuggingface [ https://huggingface.co ] :\n- The place to store, merge, and tune models endlessly.\n- THE reason we have an open source community.\n\nLlamaCPP [ https://github.com/ggml-org/llama.cpp ] :\n- The ability to compress and run models on GPU(s), CPU(s) and almost all devices.\n- Imatrix, Quantization, and other tools to tune the quants and the models.\n- Llama-Server : A cli based direct interface to run GGUF models.\n- The only tool I use to quant models.\n\nQuant-Masters: Team Mradermacher, Bartowski, and many others:\n- Quant models day and night for us all to use.\n- They are the lifeblood of open source access.\n\nMergeKit [ https://github.com/arcee-ai/mergekit ] :\n- The universal online/offline tool to merge models together and forge something new.\n- Over 20 methods to almost instantly merge model, pull them apart and put them together again.\n- The tool I have used to create over 1500 models.\n\nLmstudio [ https://lmstudio.ai/ ] :\n- The go to tool to test and run models in GGUF format.\n- The Tool I use to test/refine and evaluate new models.\n- LMStudio forum on discord; endless info and community for open source.\n\nText Generation Webui // KolboldCPP // SillyTavern:\n- Excellent tools to run GGUF models with - [ https://github.com/oobabooga/text-generation-webui ] [ https://github.com/LostRuins/koboldcpp ] .\n- Sillytavern [ https://github.com/SillyTavern/SillyTavern ] can be used with LMSTudio [ https://lmstudio.ai/ ] , TextGen [ https://github.com/oobabooga/text-generation-webui ], Kolboldcpp [ https://github.com/LostRuins/koboldcpp ], Llama-Server [part of LLAMAcpp] as a off the scale front end control system and interface to work with models.\n",
"related_quantizations": []
},
"tags": [
"gguf",
"story",
"roleplay",
"creative",
"rp",
"fantasy",
"story telling",
"32 bit upscale",
"ultra high precision",
"nsfw",
"llama",
"llama-2",
"text-generation",
"en",
"license:apache-2.0",
"endpoints_compatible",
"region:us"
],
"likes": 15,
"downloads": 492,
"gated": false,
"private": false,
"last_modified": "2025-07-28T00:28:04.000Z",
"created_at": "2024-06-17T10:35:23.000Z",
"pipeline_tag": "text-generation",
"library_name": ""
}
Source payload excerpt (from Hugging Face API)
{
"_id": "6670116b83571a7a0562f53d",
"id": "DavidAU/DarkForest-20B-V3-Ultra-Quality-GGUF",
"modelId": "DavidAU/DarkForest-20B-V3-Ultra-Quality-GGUF",
"sha": "6701368114657e60cd7c9bbea910420b58960952",
"createdAt": "2024-06-17T10:35:23.000Z",
"lastModified": "2025-07-28T00:28:04.000Z",
"author": "DavidAU",
"downloads": 492,
"likes": 15,
"gated": false,
"private": false,
"pipeline_tag": "text-generation",
"library_name": "",
"siblings_count": 14
}