GraySoft
Projects Models About FAQ Contact Download guIDE →

davidau/qwen3-4b-hivemind-instruct-neo-max-imatrix-gguf - Free GGUF Download is indexed on GraySoft with repository links, GGUF quant files, and Hugging Face metadata. This page helps you pick a local model for guIDE or other runtimes. See related models in the same shard below.

Model Intelligence Sheet

davidau/qwen3-4b-hivemind-instruct-neo-max-imatrix-gguf overview

Qwen3-4B-Hivemind-Instruct-NEO-MAX-Imatrix-GGUF The Storm is coming... 256k context, off the scale power. NEO Imatrix with MAX quants (16 bit OT all quants). This is the one that will make closed source sweat. This is the one that will make them all sweat. This is a general purpose model. BENCHMARKS: WANT POWER (4B/6B/8B) without "the nanny" ? https://huggingface.co/DavidAU/Qwen3-4B-Hivemind-Instruct-Heretic-Abliterated-Uncensored-NEO-Imatrix-GGUF https://huggingface.co/DavidAU/Qwen3-6B-Hivemind-Instruct-Heretic-Abliterated-Uncensored-NEO-Imatrix-GGUF https://huggingface.co/DavidAU/Qwen3-8B-Hivemind-Instruct-Heretic-Abliterated-Uncensored-NEO-Imatrix-GGUF --- Special Thanks: --- This was a Colab project between Nightmedia and DavidAU. Nightmedia: https://huggingface.co/nightmedia Thanks to the following model makers/tuners (models used in this project): https://huggingface.co/Gen-Verse/Qwen3-4B-RA-SFT https://huggingface.co/TeichAI/Qwen3-4B-Instruct-2507-Polaris-Alpha-Distill https://huggingface.co/TeichAI/Qwen3-4B-Thinking-2507-Gemini-2.5-Flash-Distill And of course team Qwen: https://huggingface.co/Qwen --- Settings: CHAT / ROLEPLAY and/or SMOOTHER operation of this model: In "KoboldCpp" or "oobabooga/text-generation-webui" or "Silly Tavern" ; Set the "Smoothingfactor" to 1.5 : in KoboldCpp -> Settings->Samplers->Advanced-> "SmoothF" : in text-generation-webui -> parameters -> lower right. : In Silly Tavern this is called: "Smoothing" NOTE: For "text-generation-webui" -> if using GGUFs you need to use "llamaHF" (which involves downloading some config files from the SOURCE version of this model) Source versions (and config files) of my models are here: https://huggingface.co/collections/DavidAU/d-au-source-files-for-gguf-exl2-awq-gptq-hqq-etc-etc-66b55cb8ba25f914cbf210be OTHER OPTIONS: Highest Quality Settings / Optimal Operation Guide / Parameters and Samplers This a "Class 1" model: For all settings used for this model (including specifics for its "class"), including example generation(s) and for advanced settings guide (which many times addresses any model issue(s)), including methods to improve model performance for all use case(s) as well as chat, roleplay and other use case(s) please see: [ https://huggingface.co/DavidAU/Maximizing-Model-Performance-All-Quants-Types-And-Full-Precision-by-SamplersParameters ] You can see all parameters used for generation, in addition to advanced parameters and samplers to get the most out of this model here: [ https://huggingface.co/DavidAU/Maximizing-Model-Performance-All-Quants-Types-And-Full-Precision-by-Samplers_Parameters ]

gguf256k contextQwen3All use casescreativecreative writingfiction writingplot generationsub-plot generationstory generationscene continuestorytellingfiction storyscience fictionromanceall genresstorywritingvivid prosingvivid writingfictionroleplayingbfloat16Neo Imatrix Maxtext-generationenlicense:apache-2.0endpoints_compatibleregion:usimatrix
davidau/qwen3-4b-hivemind-instruct-neo-max-imatrix-gguf visual
Downloads
326
Likes
7
Pipeline
text-generation
Library
Visibility
Public
Access
Open

Repository Files & Downloads

8 files detected
Direct downloads for all repository files
FileTypeQuantizationSizeLink
Qwen3-4B-Hivemind-Instruct-NeoMAX-D_AU-IQ2_M-imat.gguf GGUF IQ2_M 2.04 GB Download
Qwen3-4B-Hivemind-Instruct-NeoMAX-D_AU-IQ3_M-imat.gguf GGUF IQ3_M 2.41 GB Download
Qwen3-4B-Hivemind-Instruct-NeoMAX-D_AU-IQ4_XS-imat.gguf GGUF IQ4_XS 2.73 GB Download
Qwen3-4B-Hivemind-Instruct-NeoMAX-D_AU-Q4_K_M-imat.gguf GGUF Q4_K_M 2.96 GB Download
Qwen3-4B-Hivemind-Instruct-NeoMAX-D_AU-Q5_K_M-imat.gguf GGUF Q5_K_M 3.37 GB Download
Qwen3-4B-Hivemind-Instruct-NeoMAX-D_AU-Q5_K_S-imat.gguf GGUF Q5_K_S 3.31 GB Download
Qwen3-4B-Hivemind-Instruct-NeoMAX-D_AU-Q6_K-imat.gguf GGUF Q6_K 3.80 GB Download
Qwen3-4B-Hivemind-Instruct-NeoMAX-D_AU-Q8_0.gguf GGUF 4.71 GB Download

Model Details Live

Model Slug
davidau/qwen3-4b-hivemind-instruct-neo-max-imatrix-gguf
Author
DavidAU
Pipeline Task
text-generation
Library
Created
2025-12-05
Last Modified
2025-12-05
Gated
No
Private
No
HF SHA
fcdf7865bbc08e56e1f39ed40c13e65a205cd671
License
apache-2.0
Language
en
Base Model
Unknown

Metadata Inspector

Normalized metadata (stored in metadata_json)
{
  "metadata": {},
  "card_data": {
    "license": "apache-2.0",
    "language": [
      "en"
    ],
    "tags": [
      "256k context",
      "Qwen3",
      "All use cases",
      "creative",
      "creative writing",
      "fiction writing",
      "plot generation",
      "sub-plot generation",
      "fiction writing",
      "story generation",
      "scene continue",
      "storytelling",
      "fiction story",
      "science fiction",
      "romance",
      "all genres",
      "story",
      "writing",
      "vivid prosing",
      "vivid writing",
      "fiction",
      "roleplaying",
      "bfloat16",
      "Neo Imatrix Max"
    ],
    "pipeline_tag": "text-generation",
    "frontmatter": {
      "license": "apache-2.0",
      "language": [
        "en"
      ],
      "tags": [
        "256k context",
        "Qwen3",
        "All use cases",
        "creative",
        "creative writing",
        "fiction writing",
        "plot generation",
        "sub-plot generation",
        "fiction writing",
        "story generation",
        "scene continue",
        "storytelling",
        "fiction story",
        "science fiction",
        "romance",
        "all genres",
        "story",
        "writing",
        "vivid prosing",
        "vivid writing",
        "fiction",
        "roleplaying",
        "bfloat16",
        "Neo Imatrix Max"
      ],
      "pipeline_tag": "text-generation"
    },
    "hero_image_url": "hivemind2.gif",
    "summary": "Qwen3-4B-Hivemind-Instruct-NEO-MAX-Imatrix-GGUF  The Storm is coming... 256k context, off the scale power. NEO Imatrix with MAX quants (16 bit OT all quants). This is the one that will make closed source sweat. This is the one that will make them all sweat. This is a general purpose model. BENCHMARKS: `` MODEL                       arc_challenge,arc_easy,boolq,hellaswag,openbookqa,piqa,winogrande Our 4B Instruct Model       0.613,0.842,0.855,0.748,0.428,0.781,0.709 Qwen3-30B-A3B-Thinking-2507 0.421,0.448,0.682,0.635,0.402,0.771,0.669 `` WANT POWER (4B/6B/8B) without \"the nanny\" ? https://huggingface.co/DavidAU/Qwen3-4B-Hivemind-Instruct-Heretic-Abliterated-Uncensored-NEO-Imatrix-GGUF https://huggingface.co/DavidAU/Qwen3-6B-Hivemind-Instruct-Heretic-Abliterated-Uncensored-NEO-Imatrix-GGUF https://huggingface.co/DavidAU/Qwen3-8B-Hivemind-Instruct-Heretic-Abliterated-Uncensored-NEO-Imatrix-GGUF --- Special Thanks: --- This was a Colab project between Nightmedia and DavidAU. Nightmedia: https://huggingface.co/nightmedia Thanks to the following model makers/tuners (models used in this project): https://huggingface.co/Gen-Verse/Qwen3-4B-RA-SFT https://huggingface.co/TeichAI/Qwen3-4B-Instruct-2507-Polaris-Alpha-Distill https://huggingface.co/TeichAI/Qwen3-4B-Thinking-2507-Gemini-2.5-Flash-Distill And of course team Qwen: https://huggingface.co/Qwen --- Settings: CHAT / ROLEPLAY and/or SMOOTHER operation of this model: In \"KoboldCpp\" or  \"oobabooga/text-generation-webui\" or \"Silly Tavern\" ; Set the \"Smoothing_factor\" to 1.5 : in KoboldCpp -> Settings->Samplers->Advanced-> \"Smooth_F\" : in text-generation-webui -> parameters -> lower right. : In Silly Tavern this is called: \"Smoothing\" NOTE: For \"text-generation-webui\" -> if using GGUFs you need to use \"llama_HF\" (which involves downloading some config files from the SOURCE version of this model) Source versions (and config files) of my models are here: https://huggingface.co/collections/DavidAU/d-au-source-files-for-gguf-exl2-awq-gptq-hqq-etc-etc-66b55cb8ba25f914cbf210be OTHER OPTIONS: Highest Quality Settings / Optimal Operation Guide / Parameters and Samplers This a \"Class 1\" model: For all settings used for this model (including specifics for its \"class\"), including example generation(s) and for advanced settings guide (which many times addresses any model issue(s)), including methods to improve model performance for all use case(s) as well as chat, roleplay and other use case(s) please see: [ https://huggingface.co/DavidAU/Maximizing-Model-Performance-All-Quants-Types-And-Full-Precision-by-Samplers_Parameters ] You can see all parameters used for generation, in addition to advanced parameters and samplers to get the most out of this model here: [ https://huggingface.co/DavidAU/Maximizing-Model-Performance-All-Quants-Types-And-Full-Precision-by-Samplers_Parameters ]",
    "quick_links": [],
    "benchmark_table_html": "",
    "readme_markdown": "---\nlicense: apache-2.0\nlanguage:\n- en\ntags:\n- 256k context\n- Qwen3\n- All use cases\n- creative\n- creative writing\n- fiction writing\n- plot generation\n- sub-plot generation\n- fiction writing\n- story generation\n- scene continue\n- storytelling\n- fiction story\n- science fiction\n- romance\n- all genres\n- story\n- writing\n- vivid prosing\n- vivid writing\n- fiction\n- roleplaying\n- bfloat16\n- Neo Imatrix Max\npipeline_tag: text-generation\n---\n\n<H2>Qwen3-4B-Hivemind-Instruct-NEO-MAX-Imatrix-GGUF</H2>\n\n<img src=\"hivemind2.gif\" style=\"float:right; width:300px; height:300px; padding:10px;\">\n\nThe Storm is coming...\n\n256k context, off the scale power.\n\nNEO Imatrix with MAX quants (16 bit OT all quants).\n\nThis is the one that will make closed source sweat.\n\nThis is the one that will make them all sweat.\n\nThis is a general purpose model.\n\nBENCHMARKS:\n\n```\n\nMODEL                       arc_challenge,arc_easy,boolq,hellaswag,openbookqa,piqa,winogrande\n\nOur 4B Instruct Model       0.613,0.842,0.855,0.748,0.428,0.781,0.709\nQwen3-30B-A3B-Thinking-2507 0.421,0.448,0.682,0.635,0.402,0.771,0.669\n\n```\n\nWANT POWER (4B/6B/8B) without \"the nanny\" ?\n\nhttps://huggingface.co/DavidAU/Qwen3-4B-Hivemind-Instruct-Heretic-Abliterated-Uncensored-NEO-Imatrix-GGUF\n\nhttps://huggingface.co/DavidAU/Qwen3-6B-Hivemind-Instruct-Heretic-Abliterated-Uncensored-NEO-Imatrix-GGUF\n\nhttps://huggingface.co/DavidAU/Qwen3-8B-Hivemind-Instruct-Heretic-Abliterated-Uncensored-NEO-Imatrix-GGUF\n\n---\n\n<B>Special Thanks:</B>\n\n---\n\nThis was a Colab project between Nightmedia and DavidAU.\n\nNightmedia:\n\nhttps://huggingface.co/nightmedia\n\nThanks to the following model makers/tuners (models used in this project):\n\nhttps://huggingface.co/Gen-Verse/Qwen3-4B-RA-SFT\n\nhttps://huggingface.co/TeichAI/Qwen3-4B-Instruct-2507-Polaris-Alpha-Distill\n\nhttps://huggingface.co/TeichAI/Qwen3-4B-Thinking-2507-Gemini-2.5-Flash-Distill\n\nAnd of course team Qwen:\n\nhttps://huggingface.co/Qwen\n\n---\n\n<B>Settings: CHAT / ROLEPLAY and/or SMOOTHER operation of this model:</B>\n\nIn \"KoboldCpp\" or  \"oobabooga/text-generation-webui\" or \"Silly Tavern\" ;\n\nSet the \"Smoothing_factor\" to 1.5 \n\n: in KoboldCpp -> Settings->Samplers->Advanced-> \"Smooth_F\"\n\n: in text-generation-webui -> parameters -> lower right.\n\n: In Silly Tavern this is called: \"Smoothing\"\n\n\nNOTE: For \"text-generation-webui\" \n\n-> if using GGUFs you need to use \"llama_HF\" (which involves downloading some config files from the SOURCE version of this model)\n\nSource versions (and config files) of my models are here:\n\nhttps://huggingface.co/collections/DavidAU/d-au-source-files-for-gguf-exl2-awq-gptq-hqq-etc-etc-66b55cb8ba25f914cbf210be\n\nOTHER OPTIONS:\n\n- Increase rep pen to 1.1 to 1.15 (you don't need to do this if you use \"smoothing_factor\")\n\n- If the interface/program you are using to run AI MODELS supports \"Quadratic Sampling\" (\"smoothing\") just make the adjustment as noted.\n\n<B>Highest Quality Settings / Optimal Operation Guide / Parameters and Samplers</B>\n\nThis a \"Class 1\" model:\n\nFor all settings used for this model (including specifics for its \"class\"), including example generation(s) and for advanced settings guide (which many times addresses any model issue(s)), including methods to improve model performance for all use case(s) as well as chat, roleplay and other use case(s) please see:\n\n[ https://huggingface.co/DavidAU/Maximizing-Model-Performance-All-Quants-Types-And-Full-Precision-by-Samplers_Parameters ]\n\nYou can see all parameters used for generation, in addition to advanced parameters and samplers to get the most out of this model here:\n\n[ https://huggingface.co/DavidAU/Maximizing-Model-Performance-All-Quants-Types-And-Full-Precision-by-Samplers_Parameters ]\n\n",
    "related_quantizations": []
  },
  "tags": [
    "gguf",
    "256k context",
    "Qwen3",
    "All use cases",
    "creative",
    "creative writing",
    "fiction writing",
    "plot generation",
    "sub-plot generation",
    "story generation",
    "scene continue",
    "storytelling",
    "fiction story",
    "science fiction",
    "romance",
    "all genres",
    "story",
    "writing",
    "vivid prosing",
    "vivid writing",
    "fiction",
    "roleplaying",
    "bfloat16",
    "Neo Imatrix Max",
    "text-generation",
    "en",
    "license:apache-2.0",
    "endpoints_compatible",
    "region:us",
    "imatrix",
    "conversational"
  ],
  "likes": 7,
  "downloads": 326,
  "gated": false,
  "private": false,
  "last_modified": "2025-12-05T10:42:04.000Z",
  "created_at": "2025-12-05T02:35:39.000Z",
  "pipeline_tag": "text-generation",
  "library_name": ""
}
Source payload excerpt (from Hugging Face API)
{
  "_id": "693244fb8f98df18d4f84917",
  "id": "DavidAU/Qwen3-4B-Hivemind-Instruct-NEO-MAX-Imatrix-GGUF",
  "modelId": "DavidAU/Qwen3-4B-Hivemind-Instruct-NEO-MAX-Imatrix-GGUF",
  "sha": "fcdf7865bbc08e56e1f39ed40c13e65a205cd671",
  "createdAt": "2025-12-05T02:35:39.000Z",
  "lastModified": "2025-12-05T10:42:04.000Z",
  "author": "DavidAU",
  "downloads": 326,
  "likes": 7,
  "gated": false,
  "private": false,
  "pipeline_tag": "text-generation",
  "library_name": "",
  "siblings_count": 12
}