GraySoft
Projects Models Compare Cloud benchmarks FAQ Download guIDE →
Model Intelligence Sheet

ljupco/LFM2.5-8B-A1B-GGUF overview

LFM2.5 8B A1B — GGUF mixed 4 bit quant A GGUF conversion of LiquidAI's LFM2.5 8B A1B MoE, ~1B active for llama.cpp / vllm.cpp, with a mixed quantization: the b…

ggufllama.cpplfm2.5base_model:LiquidAI/LFM2.5-8B-A1Bbase_model:quantized:LiquidAI/LFM2.5-8B-A1Blicense:otherendpoints_compatibleregion:usconversational

Runs locally from ~4.45 GB disk (8 GB VRAM class GPUs with llama.cpp / guIDE).

Downloads
0
Likes
0
Pipeline
Author

Repository Files & Downloads

1 GGUF files detected
Direct downloads for local inference
FileTypeQuantizationSizeLink
LFM2.5-8B-A1B-Q4_0h.ggufGGUFQ4_0H4.45 GBDownload

Model Details

Model IDljupco/LFM2.5-8B-A1B-GGUF
Authorljupco
Pipeline
Licenseother
Base modelLiquidAI/LFM2.5-8B-A1B
Last modified2026-08-09T10:45:07.000Z

Model README

---

license: other

base_model: LiquidAI/LFM2.5-8B-A1B

quantized_by: ljupco

tags:

  • llama.cpp
  • gguf
  • lfm2.5

---

LFM2.5-8B-A1B — GGUF (mixed 4-bit quant)

A GGUF conversion of LiquidAI's LFM2.5-8B-A1B (MoE, ~1B active) for llama.cpp /

vllm.cpp, with a mixed quantization: the bulk of the weights are 4-bit (Q4_0) while

the most sensitive tensors keep a higher precision. This is one of the four models

benchmarked in the three-engine report:

Usage

llama-cli -m LFM2.5-8B-A1B-Q4_0h.gguf -p "The capital of France is" -n 64

Credits and Acknowledgements

This is a quantization of LFM2.5-8B-A1B by Liquid AI. We are deeply grateful

to Liquid AI for the LFM2.5 family, its gated-delta / shortconv architecture, and for

publishing the weights openly. This work builds directly on theirs, and we thank them

profusely. See the report above for the full acknowledgement.

Run ljupco/LFM2.5-8B-A1B-GGUF with guIDE

Download guIDE — the AI-native code editor with local LLM inference and 69 built-in tools.

Download guIDE → · Browse 524k+ models · Compare models

Source: Hugging Face · Compare models