Model Intelligence Sheet
KoarAI/LFM2.5-350M-Thinking-0004-GGUF overview
🧠LFM2.5 350M Thinking 0004 GGUF Quantized GGUF binaries of KoarAI/LFM2.5 350M Thinking 0004 for ultra fast local inference on CPU, Vulkan, and Metal via llam…
Runs locally from ~676.3 MB disk (4 GB VRAM class GPUs with llama.cpp / guIDE).
Repository Files & Downloads
1 GGUF files detected
Direct downloads for local inference
| File | Type | Quantization | Size | Link |
|---|---|---|---|---|
| LFM2.5-350M-Thinking-0004-f16.gguf | GGUF | F16 | 676.3 MB | Download |
Model Details
Model README
---
license: apache-2.0
language:
- ru
- en
tags:
- liquid-ai
- lfm2.5
- reasoning
- gguf
- llama-cpp
- ollama
---
🧠LFM2.5-350M-Thinking-0004-GGUF
Quantized GGUF binaries of KoarAI/LFM2.5-350M-Thinking-0004 for ultra-fast local inference on CPU, Vulkan, and Metal via llama.cpp.
📦 Files Available
LFM2.5-350M-Thinking-0004-f16.gguf(~709 MB) — Full Precision FP16 baseline.
🚀 Quick Start (llama.cpp)
llama-cli -m LFM2.5-350M-Thinking-0004-f16.gguf -p "<|im_start|>user\nСколько будет 17 * 19?\n<|im_end|>\n<|im_start|>assistant\n" -n 512 --temp 0.6Run KoarAI/LFM2.5-350M-Thinking-0004-GGUF with guIDE
Download guIDE — the AI-native code editor with local LLM inference and 69 built-in tools.
Source: Hugging Face · Compare models