GraySoft
Projects Models Compare Cloud benchmarks FAQ Download guIDE →
Model Intelligence Sheet

steven0226/qwen3vl-8b-chartqa-gguf overview

Qwen3 VL 8B ChartQA — GGUF Q4 K M Persistent CPU artifact converted from the pinned fine tuned merged checkpoint steven0226/qwen3vl 8b chartqa merged 16bit@519…

llama.cppggufqwen3-vlvision-languagechartqaq4-k-mimage-text-to-textdataset:HuggingFaceM4/ChartQAbase_model:steven0226/qwen3vl-8b-chartqa-merged-16bitbase_model:quantized:steven0226/qwen3vl-8b-chartqa-merged-16bitlicense:apache-2.0endpoints_compatibleregion:usconversational

Runs locally from ~717.4 MB disk (4 GB VRAM class GPUs with llama.cpp / guIDE).

Downloads
0
Likes
0
Pipeline
image-text-to-text

Repository Files & Downloads

2 GGUF files detected
Direct downloads for local inference
FileTypeQuantizationSizeLink
Qwen3VL-8B-ChartQA-Q4_K_M.ggufGGUFQ4_K_M4.68 GBDownload
mmproj-Qwen3VL-8B-ChartQA-Q8_0.ggufGGUFQ8_0717.4 MBDownload

Model Details

Model IDsteven0226/qwen3vl-8b-chartqa-gguf
Authorsteven0226
Pipelineimage-text-to-text
Licenseapache-2.0
Base modelsteven0226/qwen3vl-8b-chartqa-merged-16bit
Last modified2026-08-13T19:32:42.000Z

Model README

---

license: apache-2.0

base_model: steven0226/qwen3vl-8b-chartqa-merged-16bit

library_name: llama.cpp

pipeline_tag: image-text-to-text

tags:

  • qwen3-vl
  • vision-language
  • chartqa
  • gguf
  • llama.cpp
  • q4-k-m

datasets:

  • HuggingFaceM4/ChartQA

---

Qwen3-VL-8B ChartQA — GGUF Q4_K_M

Persistent CPU artifact converted from the pinned fine-tuned merged checkpoint

steven0226/qwen3vl-8b-chartqa-merged-16bit@519060ef43df3261e0512e5ae4c82a4d4e675f32.

Files

  • Qwen3VL-8B-ChartQA-Q4_K_M.gguf — fine-tuned language model, Q4_K_M
  • mmproj-Qwen3VL-8B-ChartQA-Q8_0.gguf — vision encoder/projector, Q8_0

Conversion used llama.cpp commit 79bba02a6741de194912d370015866414faa83ad.

conversion_metadata.json records byte sizes, digests, lineage, and a

content-free smoke-test outcome.

Verification scope

One pinned ChartQA test sample fetched at runtime passed an exact-match check.

The chart, question, label, and model output are not published. This is a smoke

verification, not a complete GGUF quality evaluation.

The formal 2,500-question quality result belongs to the separately evaluated

AWQ/vLLM artifact:

85.52%, -0.72 pp versus merged 16-bit, passing the predefined 2 pp gate.

llama.cpp

llama-server \
  --model Qwen3VL-8B-ChartQA-Q4_K_M.gguf \
  --mmproj mmproj-Qwen3VL-8B-ChartQA-Q8_0.gguf \
  --ctx-size 4096 --jinja

Publication boundary

This model repository contains weights and content-free metadata. It does not

redistribute ChartQA images, questions, labels, or raw model predictions.

Run steven0226/qwen3vl-8b-chartqa-gguf with guIDE

Download guIDE — the AI-native code editor with local LLM inference and 69 built-in tools.

Download guIDE → · Browse 524k+ models · Compare models

Source: Hugging Face · Compare models