GraySoft
Projects Models Compare Cloud benchmarks FAQ Download guIDE →
Model Intelligence Sheet

JamieBradfield/qwen3.8-9b-hermes-function-calling-v2-GGUF overview

Qwen3.8 9B Hermes FC v2 — GGUF ROCmFPX ROCmFPX quantized GGUF of the BF16 merge https://huggingface.co/JamieBradfield/qwen3.8 9b hermes function calling v2 . F…

ggufqwen3.5rocmfpxfunction-callingtext-generationbase_model:JamieBradfield/qwen3.8-9b-hermes-function-calling-v2base_model:quantized:JamieBradfield/qwen3.8-9b-hermes-function-calling-v2license:apache-2.0endpoints_compatibleregion:usconversational

Runs locally from ~4.58 GB disk (8 GB VRAM class GPUs with llama.cpp / guIDE).

Downloads
0
Likes
0
Pipeline
text-generation

Repository Files & Downloads

1 GGUF files detected
Direct downloads for local inference
FileTypeQuantizationSizeLink
qwen3.8-9b-hf-fc-v2-Q4_0_ROCMFP4_FAST.ggufGGUFQ4_0_ROCMFP4_FAST4.58 GBDownload

Model Details

Model IDJamieBradfield/qwen3.8-9b-hermes-function-calling-v2-GGUF
AuthorJamieBradfield
Pipelinetext-generation
Licenseapache-2.0
Base modelJamieBradfield/qwen3.8-9b-hermes-function-calling-v2
Last modified2026-08-28T22:31:33.000Z

Model README

---

license: apache-2.0

base_model: JamieBradfield/qwen3.8-9b-hermes-function-calling-v2

tags:

  • qwen3.5
  • gguf
  • rocmfpx
  • function-calling

pipeline_tag: text-generation

---

Qwen3.8-9B Hermes FC v2 — GGUF (ROCmFPX)

ROCmFPX quantized GGUF of the BF16 merge.

Files

| file | quant | size |

|---|---|---|

| qwen3.8-9b-hf-fc-v2-Q4_0_ROCMFP4_FAST.gguf | Q4_0_ROCMFP4_FAST | 4.92 GB |

Notes

  • ROCmFPX quants use ROCMFP4 kernels from the llama-rocmfpx fork

(AMD RDNA3 / RX 7700 XT). Not portable to CUDA or CPU-only builds without

the fork's kernels.

  • MTP (multi-token prediction) head is preserved in the GGUF: the BF16

merge drops only the vision tower.

  • Converted from the BF16 merge with convert_hf_to_gguf.py --outtype bf16,

quantized with llama-quantize Q4_0_ROCMFP4_FAST.

  • Model card and evaluation status: see the base-model repo.

Run JamieBradfield/qwen3.8-9b-hermes-function-calling-v2-GGUF with guIDE

Download guIDE — the AI-native code editor with local LLM inference and 69 built-in tools.

Download guIDE → · Browse 524k+ models · Compare models

Source: Hugging Face · Compare models