davidau/qwen3-4b-hivemind-instruct-neo-max-imatrix-gguf Q8_0 GGUF - Free GGUF Download is indexed on GraySoft with repository links, GGUF quant files, and Hugging Face metadata. This page helps you pick a local model for guIDE or other runtimes. See related models in the same shard below.
davidau/qwen3-4b-hivemind-instruct-neo-max-imatrix-gguf overview
Qwen3-4B-Hivemind-Instruct-NEO-MAX-Imatrix-GGUF The Storm is coming... 256k context, off the scale power. NEO Imatrix with MAX quants (16 bit OT all quants). This is the one that will make closed source sweat. This is the one that will make them all sweat. This is a general purpose model. BENCHMARKS: WANT POWER (4B/6B/8B) without "the nanny" ? https://huggingface.co/DavidAU/Qwen3-4B-Hivemind-Instruct-Heretic-Abliterated-Uncensored-NEO-Imatrix-GGUF https://huggingface.co/DavidAU/Qwen3-6B-Hivemind-Instruct-Heretic-Abliterated-Uncensored-NEO-Imatrix-GGUF https://huggingface.co/DavidAU/Qwen3-8B-Hivemind-Instruct-Heretic-Abliterated-Uncensored-NEO-Imatrix-GGUF --- Special Thanks: --- This was a Colab project between Nightmedia and DavidAU. Nightmedia: https://huggingface.co/nightmedia Thanks to the following model makers/tuners (models used in this project): https://huggingface.co/Gen-Verse/Qwen3-4B-RA-SFT https://huggingface.co/TeichAI/Qwen3-4B-Instruct-2507-Polaris-Alpha-Distill https://huggingface.co/TeichAI/Qwen3-4B-Thinking-2507-Gemini-2.5-Flash-Distill And of course team Qwen: https://huggingface.co/Qwen --- Settings: CHAT / ROLEPLAY and/or SMOOTHER operation of this model: In "KoboldCpp" or "oobabooga/text-generation-webui" or "Silly Tavern" ; Set the "Smoothingfactor" to 1.5 : in KoboldCpp -> Settings->Samplers->Advanced-> "SmoothF" : in text-generation-webui -> parameters -> lower right. : In Silly Tavern this is called: "Smoothing" NOTE: For "text-generation-webui" -> if using GGUFs you need to use "llamaHF" (which involves downloading some config files from the SOURCE version of this model) Source versions (and config files) of my models are here: https://huggingface.co/collections/DavidAU/d-au-source-files-for-gguf-exl2-awq-gptq-hqq-etc-etc-66b55cb8ba25f914cbf210be OTHER OPTIONS: Highest Quality Settings / Optimal Operation Guide / Parameters and Samplers This a "Class 1" model: For all settings used for this model (including specifics for its "class"), including example generation(s) and for advanced settings guide (which many times addresses any model issue(s)), including methods to improve model performance for all use case(s) as well as chat, roleplay and other use case(s) please see: [ https://huggingface.co/DavidAU/Maximizing-Model-Performance-All-Quants-Types-And-Full-Precision-by-SamplersParameters ] You can see all parameters used for generation, in addition to advanced parameters and samplers to get the most out of this model here: [ https://huggingface.co/DavidAU/Maximizing-Model-Performance-All-Quants-Types-And-Full-Precision-by-Samplers_Parameters ]
Repository Files & Downloads
| File | Type | Quantization | Size | Link |
|---|---|---|---|---|
| Qwen3-4B-Hivemind-Instruct-NeoMAX-D_AU-IQ2_M-imat.gguf | GGUF | IQ2_M | 2.04 GB | Download |
| Qwen3-4B-Hivemind-Instruct-NeoMAX-D_AU-IQ3_M-imat.gguf | GGUF | IQ3_M | 2.41 GB | Download |
| Qwen3-4B-Hivemind-Instruct-NeoMAX-D_AU-IQ4_XS-imat.gguf | GGUF | IQ4_XS | 2.73 GB | Download |
| Qwen3-4B-Hivemind-Instruct-NeoMAX-D_AU-Q4_K_M-imat.gguf | GGUF | Q4_K_M | 2.96 GB | Download |
| Qwen3-4B-Hivemind-Instruct-NeoMAX-D_AU-Q5_K_M-imat.gguf | GGUF | Q5_K_M | 3.37 GB | Download |
| Qwen3-4B-Hivemind-Instruct-NeoMAX-D_AU-Q5_K_S-imat.gguf | GGUF | Q5_K_S | 3.31 GB | Download |
| Qwen3-4B-Hivemind-Instruct-NeoMAX-D_AU-Q6_K-imat.gguf | GGUF | Q6_K | 3.80 GB | Download |
| Qwen3-4B-Hivemind-Instruct-NeoMAX-D_AU-Q8_0.gguf | GGUF | — | 4.71 GB | Download |
Model Details Live
Metadata Inspector
Normalized metadata (stored in metadata_json)
{
"metadata": {},
"card_data": {
"license": "apache-2.0",
"language": [
"en"
],
"tags": [
"256k context",
"Qwen3",
"All use cases",
"creative",
"creative writing",
"fiction writing",
"plot generation",
"sub-plot generation",
"fiction writing",
"story generation",
"scene continue",
"storytelling",
"fiction story",
"science fiction",
"romance",
"all genres",
"story",
"writing",
"vivid prosing",
"vivid writing",
"fiction",
"roleplaying",
"bfloat16",
"Neo Imatrix Max"
],
"pipeline_tag": "text-generation",
"frontmatter": {
"license": "apache-2.0",
"language": [
"en"
],
"tags": [
"256k context",
"Qwen3",
"All use cases",
"creative",
"creative writing",
"fiction writing",
"plot generation",
"sub-plot generation",
"fiction writing",
"story generation",
"scene continue",
"storytelling",
"fiction story",
"science fiction",
"romance",
"all genres",
"story",
"writing",
"vivid prosing",
"vivid writing",
"fiction",
"roleplaying",
"bfloat16",
"Neo Imatrix Max"
],
"pipeline_tag": "text-generation"
},
"hero_image_url": "hivemind2.gif",
"summary": "Qwen3-4B-Hivemind-Instruct-NEO-MAX-Imatrix-GGUF The Storm is coming... 256k context, off the scale power. NEO Imatrix with MAX quants (16 bit OT all quants). This is the one that will make closed source sweat. This is the one that will make them all sweat. This is a general purpose model. BENCHMARKS: `` MODEL arc_challenge,arc_easy,boolq,hellaswag,openbookqa,piqa,winogrande Our 4B Instruct Model 0.613,0.842,0.855,0.748,0.428,0.781,0.709 Qwen3-30B-A3B-Thinking-2507 0.421,0.448,0.682,0.635,0.402,0.771,0.669 `` WANT POWER (4B/6B/8B) without \"the nanny\" ? https://huggingface.co/DavidAU/Qwen3-4B-Hivemind-Instruct-Heretic-Abliterated-Uncensored-NEO-Imatrix-GGUF https://huggingface.co/DavidAU/Qwen3-6B-Hivemind-Instruct-Heretic-Abliterated-Uncensored-NEO-Imatrix-GGUF https://huggingface.co/DavidAU/Qwen3-8B-Hivemind-Instruct-Heretic-Abliterated-Uncensored-NEO-Imatrix-GGUF --- Special Thanks: --- This was a Colab project between Nightmedia and DavidAU. Nightmedia: https://huggingface.co/nightmedia Thanks to the following model makers/tuners (models used in this project): https://huggingface.co/Gen-Verse/Qwen3-4B-RA-SFT https://huggingface.co/TeichAI/Qwen3-4B-Instruct-2507-Polaris-Alpha-Distill https://huggingface.co/TeichAI/Qwen3-4B-Thinking-2507-Gemini-2.5-Flash-Distill And of course team Qwen: https://huggingface.co/Qwen --- Settings: CHAT / ROLEPLAY and/or SMOOTHER operation of this model: In \"KoboldCpp\" or \"oobabooga/text-generation-webui\" or \"Silly Tavern\" ; Set the \"Smoothing_factor\" to 1.5 : in KoboldCpp -> Settings->Samplers->Advanced-> \"Smooth_F\" : in text-generation-webui -> parameters -> lower right. : In Silly Tavern this is called: \"Smoothing\" NOTE: For \"text-generation-webui\" -> if using GGUFs you need to use \"llama_HF\" (which involves downloading some config files from the SOURCE version of this model) Source versions (and config files) of my models are here: https://huggingface.co/collections/DavidAU/d-au-source-files-for-gguf-exl2-awq-gptq-hqq-etc-etc-66b55cb8ba25f914cbf210be OTHER OPTIONS: Highest Quality Settings / Optimal Operation Guide / Parameters and Samplers This a \"Class 1\" model: For all settings used for this model (including specifics for its \"class\"), including example generation(s) and for advanced settings guide (which many times addresses any model issue(s)), including methods to improve model performance for all use case(s) as well as chat, roleplay and other use case(s) please see: [ https://huggingface.co/DavidAU/Maximizing-Model-Performance-All-Quants-Types-And-Full-Precision-by-Samplers_Parameters ] You can see all parameters used for generation, in addition to advanced parameters and samplers to get the most out of this model here: [ https://huggingface.co/DavidAU/Maximizing-Model-Performance-All-Quants-Types-And-Full-Precision-by-Samplers_Parameters ]",
"quick_links": [],
"benchmark_table_html": "",
"readme_markdown": "---\nlicense: apache-2.0\nlanguage:\n- en\ntags:\n- 256k context\n- Qwen3\n- All use cases\n- creative\n- creative writing\n- fiction writing\n- plot generation\n- sub-plot generation\n- fiction writing\n- story generation\n- scene continue\n- storytelling\n- fiction story\n- science fiction\n- romance\n- all genres\n- story\n- writing\n- vivid prosing\n- vivid writing\n- fiction\n- roleplaying\n- bfloat16\n- Neo Imatrix Max\npipeline_tag: text-generation\n---\n\n<H2>Qwen3-4B-Hivemind-Instruct-NEO-MAX-Imatrix-GGUF</H2>\n\n<img src=\"hivemind2.gif\" style=\"float:right; width:300px; height:300px; padding:10px;\">\n\nThe Storm is coming...\n\n256k context, off the scale power.\n\nNEO Imatrix with MAX quants (16 bit OT all quants).\n\nThis is the one that will make closed source sweat.\n\nThis is the one that will make them all sweat.\n\nThis is a general purpose model.\n\nBENCHMARKS:\n\n```\n\nMODEL arc_challenge,arc_easy,boolq,hellaswag,openbookqa,piqa,winogrande\n\nOur 4B Instruct Model 0.613,0.842,0.855,0.748,0.428,0.781,0.709\nQwen3-30B-A3B-Thinking-2507 0.421,0.448,0.682,0.635,0.402,0.771,0.669\n\n```\n\nWANT POWER (4B/6B/8B) without \"the nanny\" ?\n\nhttps://huggingface.co/DavidAU/Qwen3-4B-Hivemind-Instruct-Heretic-Abliterated-Uncensored-NEO-Imatrix-GGUF\n\nhttps://huggingface.co/DavidAU/Qwen3-6B-Hivemind-Instruct-Heretic-Abliterated-Uncensored-NEO-Imatrix-GGUF\n\nhttps://huggingface.co/DavidAU/Qwen3-8B-Hivemind-Instruct-Heretic-Abliterated-Uncensored-NEO-Imatrix-GGUF\n\n---\n\n<B>Special Thanks:</B>\n\n---\n\nThis was a Colab project between Nightmedia and DavidAU.\n\nNightmedia:\n\nhttps://huggingface.co/nightmedia\n\nThanks to the following model makers/tuners (models used in this project):\n\nhttps://huggingface.co/Gen-Verse/Qwen3-4B-RA-SFT\n\nhttps://huggingface.co/TeichAI/Qwen3-4B-Instruct-2507-Polaris-Alpha-Distill\n\nhttps://huggingface.co/TeichAI/Qwen3-4B-Thinking-2507-Gemini-2.5-Flash-Distill\n\nAnd of course team Qwen:\n\nhttps://huggingface.co/Qwen\n\n---\n\n<B>Settings: CHAT / ROLEPLAY and/or SMOOTHER operation of this model:</B>\n\nIn \"KoboldCpp\" or \"oobabooga/text-generation-webui\" or \"Silly Tavern\" ;\n\nSet the \"Smoothing_factor\" to 1.5 \n\n: in KoboldCpp -> Settings->Samplers->Advanced-> \"Smooth_F\"\n\n: in text-generation-webui -> parameters -> lower right.\n\n: In Silly Tavern this is called: \"Smoothing\"\n\n\nNOTE: For \"text-generation-webui\" \n\n-> if using GGUFs you need to use \"llama_HF\" (which involves downloading some config files from the SOURCE version of this model)\n\nSource versions (and config files) of my models are here:\n\nhttps://huggingface.co/collections/DavidAU/d-au-source-files-for-gguf-exl2-awq-gptq-hqq-etc-etc-66b55cb8ba25f914cbf210be\n\nOTHER OPTIONS:\n\n- Increase rep pen to 1.1 to 1.15 (you don't need to do this if you use \"smoothing_factor\")\n\n- If the interface/program you are using to run AI MODELS supports \"Quadratic Sampling\" (\"smoothing\") just make the adjustment as noted.\n\n<B>Highest Quality Settings / Optimal Operation Guide / Parameters and Samplers</B>\n\nThis a \"Class 1\" model:\n\nFor all settings used for this model (including specifics for its \"class\"), including example generation(s) and for advanced settings guide (which many times addresses any model issue(s)), including methods to improve model performance for all use case(s) as well as chat, roleplay and other use case(s) please see:\n\n[ https://huggingface.co/DavidAU/Maximizing-Model-Performance-All-Quants-Types-And-Full-Precision-by-Samplers_Parameters ]\n\nYou can see all parameters used for generation, in addition to advanced parameters and samplers to get the most out of this model here:\n\n[ https://huggingface.co/DavidAU/Maximizing-Model-Performance-All-Quants-Types-And-Full-Precision-by-Samplers_Parameters ]\n\n",
"related_quantizations": []
},
"tags": [
"gguf",
"256k context",
"Qwen3",
"All use cases",
"creative",
"creative writing",
"fiction writing",
"plot generation",
"sub-plot generation",
"story generation",
"scene continue",
"storytelling",
"fiction story",
"science fiction",
"romance",
"all genres",
"story",
"writing",
"vivid prosing",
"vivid writing",
"fiction",
"roleplaying",
"bfloat16",
"Neo Imatrix Max",
"text-generation",
"en",
"license:apache-2.0",
"endpoints_compatible",
"region:us",
"imatrix",
"conversational"
],
"likes": 7,
"downloads": 326,
"gated": false,
"private": false,
"last_modified": "2025-12-05T10:42:04.000Z",
"created_at": "2025-12-05T02:35:39.000Z",
"pipeline_tag": "text-generation",
"library_name": ""
}
Source payload excerpt (from Hugging Face API)
{
"_id": "693244fb8f98df18d4f84917",
"id": "DavidAU/Qwen3-4B-Hivemind-Instruct-NEO-MAX-Imatrix-GGUF",
"modelId": "DavidAU/Qwen3-4B-Hivemind-Instruct-NEO-MAX-Imatrix-GGUF",
"sha": "fcdf7865bbc08e56e1f39ed40c13e65a205cd671",
"createdAt": "2025-12-05T02:35:39.000Z",
"lastModified": "2025-12-05T10:42:04.000Z",
"author": "DavidAU",
"downloads": 326,
"likes": 7,
"gated": false,
"private": false,
"pipeline_tag": "text-generation",
"library_name": "",
"siblings_count": 12
}