llmware/qwen3-4b-instruct-gguf - Free GGUF Download is indexed on GraySoft with repository links, GGUF quant files, and Hugging Face metadata. This page helps you pick a local model for guIDE or other runtimes. See related models in the same shard below.
Model Intelligence Sheet
llmware/qwen3-4b-instruct-gguf overview
qwen3-4b-instruct-gguf is a GGUF Q4KM int4 quantized version of Qwen3-4B-Instruct, providing a very fast inference implementation, optimized for AI PCs. This is from the latest release series from Qwen, and has 'thinking' capability expressed as 'think' tokens. This model will run on an AI PC with at least 16 GB of memory. ### Model Description
Downloads
669
Likes
1
Pipeline
—
Library
—
Visibility
Public
Access
Open
Repository Files & Downloads
1 files detected
Direct downloads for all repository files
| File | Type | Quantization | Size | Link |
|---|---|---|---|---|
| Qwen3-4B-Q4_K_M.gguf | GGUF | Q4_K_M | 2.33 GB | Download |
Model Details Live
Metadata Inspector
Normalized metadata (stored in metadata_json)
{
"metadata": {},
"card_data": {
"license": "apache-2.0",
"inference": false,
"base_model": "Qwen/Qwen3-4B-Instruct",
"base_model_relation": "quantized",
"tags": [
"green",
"llmware-chat",
"p4",
"gguf",
"emerald"
],
"frontmatter": {
"license": "apache-2.0",
"inference": "false",
"base_model": "Qwen/Qwen3-4B-Instruct",
"base_model_relation": "quantized",
"tags": "[green, llmware-chat, p4, gguf,emerald]"
},
"hero_image_url": "",
"summary": "**qwen3-4b-instruct-gguf** is a GGUF Q4_K_M int4 quantized version of Qwen3-4B-Instruct, providing a very fast inference implementation, optimized for AI PCs. This is from the latest release series from Qwen, and has 'thinking' capability expressed as 'think' tokens. This model will run on an AI PC with at least 16 GB of memory. ### Model Description",
"quick_links": [],
"benchmark_table_html": "",
"readme_markdown": "---\nlicense: apache-2.0\ninference: false \nbase_model: Qwen/Qwen3-4B-Instruct\nbase_model_relation: quantized \ntags: [green, llmware-chat, p4, gguf,emerald]\n---\n\n# qwen3-4b-instruct-gguf\n\n**qwen3-4b-instruct-gguf** is a GGUF Q4_K_M int4 quantized version of [Qwen3-4B-Instruct](https://www.huggingface.co/Qwen/Qwen3-4B-Instruct), providing a very fast inference implementation, optimized for AI PCs. \n\nThis is from the latest release series from Qwen, and has 'thinking' capability expressed as 'think' tokens. \n\nThis model will run on an AI PC with at least 16 GB of memory. \n\n### Model Description\n\n- **Developed by:** Qwen\n- **Model type:** qwen3\n- **Parameters:** 4 billion \n- **Model Parent:** Qwen/Qwen3-4B-Instruct \n- **Language(s) (NLP):** English \n- **License:** Apache 2.0 \n- **Uses:** Chat, general-purpose LLM \n- **Quantization:** int4 \n \n\n## Model Card Contact \n\n[llmware on github](https://www.github.com/llmware-ai/llmware) \n\n[llmware on hf](https://www.huggingface.co/llmware) \n\n[llmware website](https://www.llmware.ai) \n",
"related_quantizations": []
},
"tags": [
"gguf",
"qwen3",
"green",
"llmware-chat",
"p4",
"emerald",
"license:apache-2.0",
"region:us",
"conversational"
],
"likes": 1,
"downloads": 669,
"gated": false,
"private": false,
"last_modified": "2025-07-05T14:02:46.000Z",
"created_at": "2025-07-05T12:59:57.000Z",
"pipeline_tag": "",
"library_name": ""
}
Source payload excerpt (from Hugging Face API)
{
"_id": "686921cd89e9ff1bf94c22e2",
"id": "llmware/qwen3-4b-instruct-gguf",
"modelId": "llmware/qwen3-4b-instruct-gguf",
"sha": "e31d1db7449ca59b20330e7f044b59b4a0ca70f9",
"createdAt": "2025-07-05T12:59:57.000Z",
"lastModified": "2025-07-05T14:02:46.000Z",
"author": "llmware",
"downloads": 669,
"likes": 1,
"gated": false,
"private": false,
"pipeline_tag": "",
"library_name": "",
"siblings_count": 5
}