ggml-org/qwen3-omni-30b-a3b-thinking-gguf Q4_K_M GGUF - Free GGUF Download is indexed on GraySoft with repository links, GGUF quant files, and Hugging Face metadata. This page helps you pick a local model for guIDE or other runtimes. See related models in the same shard below.
Model Intelligence Sheet
ggml-org/qwen3-omni-30b-a3b-thinking-gguf overview
This model is converted from Qwen/Qwen3-Omni-30B-A3B-Thinking to GGUF using converthfto_gguf.py To use it:
Downloads
1,628
Likes
3
Pipeline
—
Library
—
Visibility
Public
Access
Open
Repository Files & Downloads
5 files detected
Direct downloads for all repository files
| File | Type | Quantization | Size | Link |
|---|---|---|---|---|
| Qwen3-Omni-30B-A3B-Thinking-Q4_K_M.gguf | GGUF | Q4_K_M | 17.28 GB | Download |
| Qwen3-Omni-30B-A3B-Thinking-Q8_0.gguf | GGUF | — | 30.25 GB | Download |
| Qwen3-Omni-30B-A3B-Thinking-bf16.gguf | GGUF | BF16 | 56.90 GB | Download |
| mmproj-Qwen3-Omni-30B-A3B-Thinking-Q8_0.gguf | GGUF | — | 1.23 GB | Download |
| mmproj-Qwen3-Omni-30B-A3B-Thinking-bf16.gguf | GGUF | BF16 | 2.06 GB | Download |
Model Details Live
Metadata Inspector
Normalized metadata (stored in metadata_json)
{
"metadata": {},
"card_data": {
"base_model": "Qwen/Qwen3-Omni-30B-A3B-Thinking",
"frontmatter": {
"base_model": "Qwen/Qwen3-Omni-30B-A3B-Thinking"
},
"hero_image_url": "",
"summary": "This model is converted from Qwen/Qwen3-Omni-30B-A3B-Thinking to GGUF using convert_hf_to_gguf.py To use it: `` llama-server -hf ggml-org/Qwen3-Omni-30B-A3B-Thinking-GGUF ``",
"quick_links": [],
"benchmark_table_html": "",
"readme_markdown": "---\nbase_model: Qwen/Qwen3-Omni-30B-A3B-Thinking\n---\n\n# Qwen3-Omni-30B-A3B-Thinking-GGUF\n\nThis model is converted from [Qwen/Qwen3-Omni-30B-A3B-Thinking](https://huggingface.co/Qwen/Qwen3-Omni-30B-A3B-Thinking) to GGUF using `convert_hf_to_gguf.py`\n\nTo use it:\n\n```\nllama-server -hf ggml-org/Qwen3-Omni-30B-A3B-Thinking-GGUF\n```\n\n",
"related_quantizations": []
},
"tags": [
"gguf",
"base_model:Qwen/Qwen3-Omni-30B-A3B-Thinking",
"base_model:quantized:Qwen/Qwen3-Omni-30B-A3B-Thinking",
"endpoints_compatible",
"region:us",
"conversational"
],
"likes": 3,
"downloads": 1628,
"gated": false,
"private": false,
"last_modified": "2026-04-13T00:23:14.000Z",
"created_at": "2026-04-13T00:20:21.000Z",
"pipeline_tag": "",
"library_name": ""
}
Source payload excerpt (from Hugging Face API)
{
"_id": "69dc36c557d19422dd4fae20",
"id": "ggml-org/Qwen3-Omni-30B-A3B-Thinking-GGUF",
"modelId": "ggml-org/Qwen3-Omni-30B-A3B-Thinking-GGUF",
"sha": "6807f126832efa5bce2969fec85a80594e21df9d",
"createdAt": "2026-04-13T00:20:21.000Z",
"lastModified": "2026-04-13T00:23:14.000Z",
"author": "ggml-org",
"downloads": 1628,
"likes": 3,
"gated": false,
"private": false,
"pipeline_tag": "",
"library_name": "",
"siblings_count": 7
}