GraySoft
Projects Models About FAQ Contact Download guIDE →

ggml-org/qwen3-omni-30b-a3b-thinking-gguf BF16 GGUF - Free GGUF Download is indexed on GraySoft with repository links, GGUF quant files, and Hugging Face metadata. This page helps you pick a local model for guIDE or other runtimes. See related models in the same shard below.

Model Intelligence Sheet

ggml-org/qwen3-omni-30b-a3b-thinking-gguf overview

This model is converted from Qwen/Qwen3-Omni-30B-A3B-Thinking to GGUF using converthfto_gguf.py To use it:

ggufbase_model:Qwen/Qwen3-Omni-30B-A3B-Thinkingbase_model:quantized:Qwen/Qwen3-Omni-30B-A3B-Thinkingendpoints_compatibleregion:usconversational
ggml-org/qwen3-omni-30b-a3b-thinking-gguf visual
Downloads
1,628
Likes
3
Pipeline
Library
Visibility
Public
Access
Open

Repository Files & Downloads

5 files detected
Direct downloads for all repository files
FileTypeQuantizationSizeLink
Qwen3-Omni-30B-A3B-Thinking-Q4_K_M.gguf GGUF Q4_K_M 17.28 GB Download
Qwen3-Omni-30B-A3B-Thinking-Q8_0.gguf GGUF 30.25 GB Download
Qwen3-Omni-30B-A3B-Thinking-bf16.gguf GGUF BF16 56.90 GB Download
mmproj-Qwen3-Omni-30B-A3B-Thinking-Q8_0.gguf GGUF 1.23 GB Download
mmproj-Qwen3-Omni-30B-A3B-Thinking-bf16.gguf GGUF BF16 2.06 GB Download

Model Details Live

Model Slug
ggml-org/qwen3-omni-30b-a3b-thinking-gguf
Author
ggml-org
Pipeline Task
Library
Created
2026-04-13
Last Modified
2026-04-13
Gated
No
Private
No
HF SHA
6807f126832efa5bce2969fec85a80594e21df9d
License
Unknown
Language
Unknown
Base Model
Qwen/Qwen3-Omni-30B-A3B-Thinking

Metadata Inspector

Normalized metadata (stored in metadata_json)
{
  "metadata": {},
  "card_data": {
    "base_model": "Qwen/Qwen3-Omni-30B-A3B-Thinking",
    "frontmatter": {
      "base_model": "Qwen/Qwen3-Omni-30B-A3B-Thinking"
    },
    "hero_image_url": "",
    "summary": "This model is converted from Qwen/Qwen3-Omni-30B-A3B-Thinking to GGUF using convert_hf_to_gguf.py To use it: `` llama-server -hf ggml-org/Qwen3-Omni-30B-A3B-Thinking-GGUF ``",
    "quick_links": [],
    "benchmark_table_html": "",
    "readme_markdown": "---\nbase_model: Qwen/Qwen3-Omni-30B-A3B-Thinking\n---\n\n# Qwen3-Omni-30B-A3B-Thinking-GGUF\n\nThis model is converted from [Qwen/Qwen3-Omni-30B-A3B-Thinking](https://huggingface.co/Qwen/Qwen3-Omni-30B-A3B-Thinking) to GGUF using `convert_hf_to_gguf.py`\n\nTo use it:\n\n```\nllama-server -hf ggml-org/Qwen3-Omni-30B-A3B-Thinking-GGUF\n```\n\n",
    "related_quantizations": []
  },
  "tags": [
    "gguf",
    "base_model:Qwen/Qwen3-Omni-30B-A3B-Thinking",
    "base_model:quantized:Qwen/Qwen3-Omni-30B-A3B-Thinking",
    "endpoints_compatible",
    "region:us",
    "conversational"
  ],
  "likes": 3,
  "downloads": 1628,
  "gated": false,
  "private": false,
  "last_modified": "2026-04-13T00:23:14.000Z",
  "created_at": "2026-04-13T00:20:21.000Z",
  "pipeline_tag": "",
  "library_name": ""
}
Source payload excerpt (from Hugging Face API)
{
  "_id": "69dc36c557d19422dd4fae20",
  "id": "ggml-org/Qwen3-Omni-30B-A3B-Thinking-GGUF",
  "modelId": "ggml-org/Qwen3-Omni-30B-A3B-Thinking-GGUF",
  "sha": "6807f126832efa5bce2969fec85a80594e21df9d",
  "createdAt": "2026-04-13T00:20:21.000Z",
  "lastModified": "2026-04-13T00:23:14.000Z",
  "author": "ggml-org",
  "downloads": 1628,
  "likes": 3,
  "gated": false,
  "private": false,
  "pipeline_tag": "",
  "library_name": "",
  "siblings_count": 7
}