GraySoft
Projects Models Compare Cloud benchmarks FAQ Download guIDE →
Model Intelligence Sheet

prithivMLmods/LiquidAI-LFM2.5-230M-GGUF overview

LiquidAI LFM2.5 230M GGUF LFM2.5 230M https://huggingface.co/LiquidAI/LFM2.5 230M is Liquid AI's https://huggingface.co/LiquidAI most compact hybrid model to d…

transformersggufliquidlfm2.5edgetext-generationenarzhfrdejakoesptitbase_model:LiquidAI/LFM2.5-230Mbase_model:quantized:LiquidAI/LFM2.5-230Mlicense:otherendpoints_compatibleregion:usconversational

Runs locally from ~110.3 MB disk (4 GB VRAM class GPUs with llama.cpp / guIDE).

Downloads
0
Likes
1
Pipeline
text-generation

Repository Files & Downloads

15 GGUF files detected
Direct downloads for local inference
FileTypeQuantizationSizeLink
LFM2.5-230M.BF16.ggufGGUFGGUF440.5 MBDownload
LFM2.5-230M.F16.ggufGGUFGGUF440.5 MBDownload
LFM2.5-230M.F32.ggufGGUFGGUF878.5 MBDownload
LFM2.5-230M.Q2_K.ggufGGUFGGUF110.3 MBDownload
LFM2.5-230M.Q3_K_L.ggufGGUFGGUF132.9 MBDownload
LFM2.5-230M.Q3_K_M.ggufGGUFGGUF127.6 MBDownload
LFM2.5-230M.Q3_K_S.ggufGGUFGGUF121.6 MBDownload
LFM2.5-230M.Q4_0.ggufGGUFGGUF142.2 MBDownload
LFM2.5-230M.Q4_K_M.ggufGGUFGGUF146.3 MBDownload
LFM2.5-230M.Q4_K_S.ggufGGUFGGUF142.7 MBDownload
LFM2.5-230M.Q5_0.ggufGGUFGGUF161.5 MBDownload
LFM2.5-230M.Q5_K_M.ggufGGUFGGUF163.7 MBDownload
LFM2.5-230M.Q5_K_S.ggufGGUFGGUF161.5 MBDownload
LFM2.5-230M.Q6_K.ggufGGUFGGUF182.1 MBDownload
LFM2.5-230M.Q8_0.ggufGGUFGGUF235.2 MBDownload

Model Details

Model IDprithivMLmods/LiquidAI-LFM2.5-230M-GGUF
AuthorprithivMLmods
Pipelinetext-generation
Licenseother
Base modelLiquidAI/LFM2.5-230M
Last modified2026-06-26T01:14:26.000Z

Model README

---

license: other

license_name: lfm1.0

license_link: LICENSE

base_model:

  • LiquidAI/LFM2.5-230M

language:

  • en
  • ar
  • zh
  • fr
  • de
  • ja
  • ko
  • es
  • pt
  • it

pipeline_tag: text-generation

tags:

  • liquid
  • lfm2.5
  • edge

library_name: transformers

---

LiquidAI-LFM2.5-230M-GGUF

> LFM2.5-230M is Liquid AI's most compact hybrid model to date, a 230-million-parameter, general-purpose instruction-tuned text model built on the LFM2 architecture with extended pre-training (19T tokens) and reinforcement learning, designed specifically for on-device deployment in the tightest memory and compute budgets. Its 14-layer architecture combines 8 double-gated LIV convolution blocks with 6 GQA blocks, supports a 32,768-token context window across 10 languages, and was distilled from the larger LFM2.5-350M before being refined with multi-stage reinforcement learning, making it well-suited for agentic tasks like tool use and data extraction rather than reasoning-heavy workloads such as advanced math, code generation, or creative writing. It delivers strong edge inference throughput — 213 tok/s decode speed on a Galaxy S25 Ultra and 42 tok/s on a Raspberry Pi 5 — and despite its tiny size, outperforms similarly-scaled competitors like Granite 4.0-350M and LFM2-350M on benchmarks including IFEval (71.71), BFCLv3 (43.26), and Multi-IF (37.70), while supporting native function calling via Pythonic tool calls and ChatML-style chat templates, with deployment options spanning Transformers, vLLM, llama.cpp (GGUF), ONNX, and MLX formats.

Model Files

File Name | Quant Type | File Size | File Link |

|-----------|------------|-----------|-----------|

| LFM2.5-230M.BF16.gguf | BF16 | 462 MB | Download |

| LFM2.5-230M.F16.gguf | F16 | 462 MB | Download |

| LFM2.5-230M.F32.gguf | F32 | 921 MB | Download |

| LFM2.5-230M.Q2_K.gguf | Q2_K | 116 MB | Download |

| LFM2.5-230M.Q3_K_L.gguf | Q3_K_L | 139 MB | Download |

| LFM2.5-230M.Q3_K_M.gguf | Q3_K_M | 134 MB | Download |

| LFM2.5-230M.Q3_K_S.gguf | Q3_K_S | 127 MB | Download |

| LFM2.5-230M.Q4_0.gguf | Q4_0 | 149 MB | Download |

| LFM2.5-230M.Q4_K_M.gguf | Q4_K_M | 153 MB | Download |

| LFM2.5-230M.Q4_K_S.gguf | Q4_K_S | 150 MB | Download |

| LFM2.5-230M.Q5_0.gguf | Q5_0 | 169 MB | Download |

| LFM2.5-230M.Q5_K_M.gguf | Q5_K_M | 172 MB | Download |

| LFM2.5-230M.Q5_K_S.gguf | Q5_K_S | 169 MB | Download |

| LFM2.5-230M.Q6_K.gguf | Q6_K | 191 MB | Download |

| LFM2.5-230M.Q8_0.gguf | Q8_0 | 247 MB | Download |

llama.cpp

LLM inference in C/C++ — https://github.com/ggml-org/llama.cpp

Run prithivMLmods/LiquidAI-LFM2.5-230M-GGUF with guIDE

Download guIDE — the AI-native code editor with local LLM inference and 69 built-in tools.

Download guIDE → · Browse 524k+ models · Compare models

Source: Hugging Face · Compare models