GraySoft
Projects Models About FAQ Contact Download guIDE →

llmware/qwen3-4b-instruct-gguf - Free GGUF Download is indexed on GraySoft with repository links, GGUF quant files, and Hugging Face metadata. This page helps you pick a local model for guIDE or other runtimes. See related models in the same shard below.

Model Intelligence Sheet

llmware/qwen3-4b-instruct-gguf overview

qwen3-4b-instruct-gguf is a GGUF Q4KM int4 quantized version of Qwen3-4B-Instruct, providing a very fast inference implementation, optimized for AI PCs. This is from the latest release series from Qwen, and has 'thinking' capability expressed as 'think' tokens. This model will run on an AI PC with at least 16 GB of memory. ### Model Description

ggufqwen3greenllmware-chatp4emeraldlicense:apache-2.0region:usconversational
llmware/qwen3-4b-instruct-gguf visual
Downloads
669
Likes
1
Pipeline
Library
Visibility
Public
Access
Open

Repository Files & Downloads

1 files detected
Direct downloads for all repository files
FileTypeQuantizationSizeLink
Qwen3-4B-Q4_K_M.gguf GGUF Q4_K_M 2.33 GB Download

Model Details Live

Model Slug
llmware/qwen3-4b-instruct-gguf
Author
llmware
Pipeline Task
Library
Created
2025-07-05
Last Modified
2025-07-05
Gated
No
Private
No
HF SHA
e31d1db7449ca59b20330e7f044b59b4a0ca70f9
License
apache-2.0
Language
Unknown
Base Model
Qwen/Qwen3-4B-Instruct

Metadata Inspector

Normalized metadata (stored in metadata_json)
{
  "metadata": {},
  "card_data": {
    "license": "apache-2.0",
    "inference": false,
    "base_model": "Qwen/Qwen3-4B-Instruct",
    "base_model_relation": "quantized",
    "tags": [
      "green",
      "llmware-chat",
      "p4",
      "gguf",
      "emerald"
    ],
    "frontmatter": {
      "license": "apache-2.0",
      "inference": "false",
      "base_model": "Qwen/Qwen3-4B-Instruct",
      "base_model_relation": "quantized",
      "tags": "[green, llmware-chat, p4, gguf,emerald]"
    },
    "hero_image_url": "",
    "summary": "**qwen3-4b-instruct-gguf** is a GGUF Q4_K_M int4 quantized version of Qwen3-4B-Instruct, providing a very fast inference implementation, optimized for AI PCs. This is from the latest release series from Qwen, and has 'thinking' capability expressed as 'think' tokens. This model will run on an AI PC with at least 16 GB of memory. ### Model Description",
    "quick_links": [],
    "benchmark_table_html": "",
    "readme_markdown": "---\nlicense: apache-2.0\ninference: false \nbase_model: Qwen/Qwen3-4B-Instruct\nbase_model_relation: quantized \ntags: [green, llmware-chat, p4, gguf,emerald]\n---\n\n# qwen3-4b-instruct-gguf\n\n**qwen3-4b-instruct-gguf** is a GGUF Q4_K_M int4 quantized version of [Qwen3-4B-Instruct](https://www.huggingface.co/Qwen/Qwen3-4B-Instruct), providing a very fast inference implementation, optimized for AI PCs.    \n\nThis is from the latest release series from Qwen, and has 'thinking' capability expressed as 'think' tokens.  \n\nThis model will run on an AI PC with at least 16 GB of memory.   \n\n### Model Description\n\n- **Developed by:** Qwen\n- **Model type:** qwen3\n- **Parameters:** 4 billion  \n- **Model Parent:** Qwen/Qwen3-4B-Instruct  \n- **Language(s) (NLP):** English  \n- **License:** Apache 2.0  \n- **Uses:** Chat, general-purpose LLM  \n- **Quantization:** int4  \n  \n\n## Model Card Contact  \n\n[llmware on github](https://www.github.com/llmware-ai/llmware) \n\n[llmware on hf](https://www.huggingface.co/llmware)  \n\n[llmware website](https://www.llmware.ai)  \n",
    "related_quantizations": []
  },
  "tags": [
    "gguf",
    "qwen3",
    "green",
    "llmware-chat",
    "p4",
    "emerald",
    "license:apache-2.0",
    "region:us",
    "conversational"
  ],
  "likes": 1,
  "downloads": 669,
  "gated": false,
  "private": false,
  "last_modified": "2025-07-05T14:02:46.000Z",
  "created_at": "2025-07-05T12:59:57.000Z",
  "pipeline_tag": "",
  "library_name": ""
}
Source payload excerpt (from Hugging Face API)
{
  "_id": "686921cd89e9ff1bf94c22e2",
  "id": "llmware/qwen3-4b-instruct-gguf",
  "modelId": "llmware/qwen3-4b-instruct-gguf",
  "sha": "e31d1db7449ca59b20330e7f044b59b4a0ca70f9",
  "createdAt": "2025-07-05T12:59:57.000Z",
  "lastModified": "2025-07-05T14:02:46.000Z",
  "author": "llmware",
  "downloads": 669,
  "likes": 1,
  "gated": false,
  "private": false,
  "pipeline_tag": "",
  "library_name": "",
  "siblings_count": 5
}