GraySoft
Projects Models About FAQ Contact Download guIDE →

aessedai/qwen3.5-35b-a3b-gguf BF16 GGUF - Free GGUF Download is indexed on GraySoft with repository links, GGUF quant files, and Hugging Face metadata. This page helps you pick a local model for guIDE or other runtimes. See related models in the same shard below.

Model Intelligence Sheet

aessedai/qwen3.5-35b-a3b-gguf overview

Comprehensive model page for aessedai/qwen3.5-35b-a3b-gguf

ggufbase_model:Qwen/Qwen3.5-35B-A3Bbase_model:quantized:Qwen/Qwen3.5-35B-A3Bendpoints_compatibleregion:usimatrixconversational
aessedai/qwen3.5-35b-a3b-gguf visual
Downloads
9,926
Likes
85
Pipeline
Library
Visibility
Public
Access
Open

Repository Files & Downloads

13 files detected
Direct downloads for all repository files
FileTypeQuantizationSizeLink
Qwen3.5-35B-A3B-IQ3_S-00001-of-00002.gguf GGUF IQ3_S 10.44 MB Download
Qwen3.5-35B-A3B-IQ3_S-00002-of-00002.gguf GGUF IQ3_S 12.65 GB Download
Qwen3.5-35B-A3B-IQ4_XS-00001-of-00002.gguf GGUF IQ4_XS 10.44 MB Download
Qwen3.5-35B-A3B-IQ4_XS-00002-of-00002.gguf GGUF IQ4_XS 16.40 GB Download
Qwen3.5-35B-A3B-Q4_K_M-00001-of-00002.gguf GGUF Q4_K_M 10.44 MB Download
Qwen3.5-35B-A3B-Q4_K_M-00002-of-00002.gguf GGUF Q4_K_M 20.62 GB Download
Qwen3.5-35B-A3B-Q5_K_M-00001-of-00002.gguf GGUF Q5_K_M 10.44 MB Download
Qwen3.5-35B-A3B-Q5_K_M-00002-of-00002.gguf GGUF Q5_K_M 24.45 GB Download
imatrix.gguf GGUF 103.27 MB Download
mmproj-Qwen3.5-35B-A3B-BF16.gguf GGUF BF16 861.00 MB Download
mmproj-Qwen3.5-35B-A3B-F16.gguf GGUF F16 857.62 MB Download
mmproj-Qwen3.5-35B-A3B-F32.gguf GGUF F32 1.66 GB Download
mmproj-Qwen3.5-35B-A3B-Q8_0.gguf GGUF 585.74 MB Download

Model Details Live

Model Slug
aessedai/qwen3.5-35b-a3b-gguf
Author
AesSedai
Pipeline Task
Library
Created
2026-02-24
Last Modified
2026-03-10
Gated
No
Private
No
HF SHA
97e7b24cd5e2cd72525cbe793d8d29c0b0f9d842
License
Unknown
Language
Unknown
Base Model
Qwen/Qwen3.5-35B-A3B

Metadata Inspector

Normalized metadata (stored in metadata_json)
{
  "metadata": {},
  "card_data": {
    "base_model": [
      "Qwen/Qwen3.5-35B-A3B"
    ],
    "frontmatter": {
      "base_model": [
        "Qwen/Qwen3.5-35B-A3B"
      ]
    },
    "hero_image_url": "kld_data/01_kld_vs_filesize.png \"Chart showing Pareto KLD analysis of quants\"",
    "summary": "",
    "quick_links": [],
    "benchmark_table_html": "",
    "readme_markdown": "---\nbase_model:\n- Qwen/Qwen3.5-35B-A3B\n---\n\n## Updates\n### 3/10/2026\nI've uploaded new quants using the new fused Up + Gate conversion, this offers up to a +10% boost in prompt processing speed from my testing.\n\n## Description\n\nThis repo contains specialized MoE-quants for Qwen3.5-35B-A3B. The idea being that given the huge size of the FFN tensors compared to the rest of the tensors in the model, it should be possible to achieve a better quality while keeping the overall size of the entire model smaller compared to a similar naive quantization. To that end, the quantization type default is kept in high quality and the FFN UP + FFN GATE tensors are quanted down along with the FFN DOWN tensors.\n\n| Quant | Size | Mixture | PPL | 1-(Mean PPL(Q)/PPL(base)) | KLD |\n| :--------- | :--------- | :------- | :------- | :------- | :------- |\n| Q5_K_M | 24.45 GiB (6.06 BPW) | Q8_0 / Q5_K / Q5_K / Q6_K | 6.534165 ± 0.041552 | -0.0182% | 0.006174 ± 0.000108 |\n| Q4_K_M | 20.62 GiB (5.11 BPW) | Q8_0 / Q4_K / Q4_K / Q5_K | 6.564172 ± 0.041821 | +0.4409% | 0.010121 ± 0.000122 |\n| IQ4_XS | 16.40 GiB (4.06 BPW) | Q8_0 / IQ3_S / IQ3_S / IQ4_XS | 6.632992 ± 0.042286 | +1.4940% | 0.024130 ± 0.000250 |\n| IQ3_S | 12.65 GiB (3.14 BPW) | Q8_0 / IQ2_S / IQ2_S / IQ3_S | 6.920534 ± 0.044619 | +5.8937% | 0.061926 ± 0.000402 |\n\n![kld_graph](kld_data/01_kld_vs_filesize.png \"Chart showing Pareto KLD analysis of quants\")\n![ppl_graph](kld_data/02_ppl_vs_filesize.png \"Chart showing Pareto PPL analysis of quants\")",
    "related_quantizations": []
  },
  "tags": [
    "gguf",
    "base_model:Qwen/Qwen3.5-35B-A3B",
    "base_model:quantized:Qwen/Qwen3.5-35B-A3B",
    "endpoints_compatible",
    "region:us",
    "imatrix",
    "conversational"
  ],
  "likes": 85,
  "downloads": 9926,
  "gated": false,
  "private": false,
  "last_modified": "2026-03-10T19:40:46.000Z",
  "created_at": "2026-02-24T23:59:40.000Z",
  "pipeline_tag": "",
  "library_name": ""
}
Source payload excerpt (from Hugging Face API)
{
  "_id": "699e3b6c22ad41f2c2756fa2",
  "id": "AesSedai/Qwen3.5-35B-A3B-GGUF",
  "modelId": "AesSedai/Qwen3.5-35B-A3B-GGUF",
  "sha": "97e7b24cd5e2cd72525cbe793d8d29c0b0f9d842",
  "createdAt": "2026-02-24T23:59:40.000Z",
  "lastModified": "2026-03-10T19:40:46.000Z",
  "author": "AesSedai",
  "downloads": 9926,
  "likes": 85,
  "gated": false,
  "private": false,
  "pipeline_tag": "",
  "library_name": "",
  "siblings_count": 23
}