GraySoft
Projects Models Compare Cloud benchmarks FAQ Download guIDE →
Model Intelligence Sheet

Koshkasa/Vortex5_Phoenix-X-26B-A4B-MXFP4_MOE-GGUF overview

What's that? MXFP4 MOE quantization of Vortex5/Phoenix X 26B A4B https://huggingface.co/Vortex5/Phoenix X 26B A4B with more aggressive compression than standar…

llama.cppggufquantizedroleplaymixed precisionmxfp4text-generationbase_model:Vortex5/Phoenix-X-26B-A4Bbase_model:quantized:Vortex5/Phoenix-X-26B-A4Blicense:apache-2.0endpoints_compatibleregion:usimatrixconversational

Runs locally from ~13.22 GB disk (16 GB VRAM class GPUs with llama.cpp / guIDE).

Downloads
0
Likes
1
Pipeline
text-generation
Author

Repository Files & Downloads

1 GGUF files detected
Direct downloads for local inference
FileTypeQuantizationSizeLink
Phoenix-X-26B-A4B-MXFP4-MOE.ggufGGUFGGUF13.22 GBDownload

Model Details

Model IDKoshkasa/Vortex5_Phoenix-X-26B-A4B-MXFP4_MOE-GGUF
AuthorKoshkasa
Pipelinetext-generation
Licenseapache-2.0
Base modelVortex5/Phoenix-X-26B-A4B
Last modified2026-08-11T20:22:55.000Z

Model README

---

license: apache-2.0

base_model:

  • Vortex5/Phoenix-X-26B-A4B

library_name: llama.cpp

pipeline_tag: text-generation

tags:

  • gguf
  • quantized
  • llama.cpp
  • roleplay
  • mixed precision
  • mxfp4

quantized_by: Koshkasa

base_model_relation: quantized

---

What's that?

MXFP4_MOE quantization of Vortex5/Phoenix-X-26B-A4B with more aggressive compression than standard mxfp4_moe. Local attention at IQ4_XS, global attention at Q5_K, attn_output in global attention layers in Q6_K.

imatrix generated by alexokita

Disclosure

My only contribution is compute. This is not my merge. Have fun.

Model card incomplete. Tests and comparisons may be uploaded at a later date.

Run Koshkasa/Vortex5_Phoenix-X-26B-A4B-MXFP4_MOE-GGUF with guIDE

Download guIDE — the AI-native code editor with local LLM inference and 69 built-in tools.

Download guIDE → · Browse 524k+ models · Compare models

Source: Hugging Face · Compare models