GraySoft
Projects Models About FAQ Contact Download guIDE →

liquidai/lfm2-24b-a2b-gguf Q8_0 GGUF - Free GGUF Download is indexed on GraySoft with repository links, GGUF quant files, and Hugging Face metadata. This page helps you pick a local model for guIDE or other runtimes. See related models in the same shard below.

Model Intelligence Sheet

liquidai/lfm2-24b-a2b-gguf overview

LFM2 is a family of hybrid models designed for on-device deployment. LFM2-24B-A2B is the largest model in the family, scaling the architecture to 24 billion parameters while keeping inference efficient. !image Find more information about LFM2-24B-A2B in our blog post.

transformersggufliquidlfm2edgemoellama.cpptext-generationenarzhfrdejakoesptbase_model:LiquidAI/LFM2-24B-A2Bbase_model:quantized:LiquidAI/LFM2-24B-A2Blicense:otherendpoints_compatibleregion:usconversational
liquidai/lfm2-24b-a2b-gguf visual
Downloads
56,748
Likes
122
Pipeline
text-generation
Library
transformers
Visibility
Public
Access
Open

Repository Files & Downloads

7 files detected
Direct downloads for all repository files
FileTypeQuantizationSizeLink
LFM2-24B-A2B-BF16.gguf GGUF BF16 44.42 GB Download
LFM2-24B-A2B-F16.gguf GGUF F16 44.42 GB Download
LFM2-24B-A2B-Q4_0.gguf GGUF 12.54 GB Download
LFM2-24B-A2B-Q4_K_M.gguf GGUF Q4_K_M 13.43 GB Download
LFM2-24B-A2B-Q5_K_M.gguf GGUF Q5_K_M 15.76 GB Download
LFM2-24B-A2B-Q6_K.gguf GGUF Q6_K 18.23 GB Download
LFM2-24B-A2B-Q8_0.gguf GGUF 23.61 GB Download

Model Details Live

Model Slug
liquidai/lfm2-24b-a2b-gguf
Author
LiquidAI
Pipeline Task
text-generation
Library
transformers
Created
2026-02-17
Last Modified
2026-03-30
Gated
No
Private
No
HF SHA
35b3ac71cf7ddf668d5869843292a8945c399dcd
License
other
Language
en, ar, zh, fr, de, ja, ko, es, pt
Base Model
LiquidAI/LFM2-24B-A2B

Metadata Inspector

Normalized metadata (stored in metadata_json)
{
  "metadata": {},
  "card_data": {
    "library_name": "transformers",
    "license": "other",
    "license_name": "lfm1.0",
    "license_link": "LICENSE",
    "language": [
      "en",
      "ar",
      "zh",
      "fr",
      "de",
      "ja",
      "ko",
      "es",
      "pt"
    ],
    "pipeline_tag": "text-generation",
    "tags": [
      "liquid",
      "lfm2",
      "edge",
      "moe",
      "llama.cpp",
      "gguf"
    ],
    "base_model": [
      "LiquidAI/LFM2-24B-A2B"
    ],
    "frontmatter": {
      "library_name": "transformers",
      "license": "other",
      "license_name": "lfm1.0",
      "license_link": "LICENSE",
      "language": [
        "en",
        "ar",
        "zh",
        "fr",
        "de",
        "ja",
        "ko",
        "es",
        "pt"
      ],
      "pipeline_tag": "text-generation",
      "tags": [
        "liquid",
        "lfm2",
        "edge",
        "moe",
        "llama.cpp",
        "gguf"
      ],
      "base_model": [
        "LiquidAI/LFM2-24B-A2B"
      ]
    },
    "hero_image_url": "https://cdn-uploads.huggingface.co/production/uploads/61b8e2ba285851687028d395/2b08LKpev0DNEk6DlnWkY.png",
    "summary": "LFM2 is a family of hybrid models designed for on-device deployment. LFM2-24B-A2B is the largest model in the family, scaling the architecture to 24 billion parameters while keeping inference efficient. !image Find more information about LFM2-24B-A2B in our blog post.",
    "quick_links": [],
    "benchmark_table_html": "",
    "readme_markdown": "---\nlibrary_name: transformers\nlicense: other\nlicense_name: lfm1.0\nlicense_link: LICENSE\nlanguage:\n- en\n- ar\n- zh\n- fr\n- de\n- ja\n- ko\n- es\n- pt\npipeline_tag: text-generation\ntags:\n- liquid\n- lfm2\n- edge\n- moe\n- llama.cpp\n- gguf\nbase_model:\n- LiquidAI/LFM2-24B-A2B\n---\n\n<center>\n<div style=\"text-align: center;\">\n  <img\n    src=\"https://cdn-uploads.huggingface.co/production/uploads/61b8e2ba285851687028d395/2b08LKpev0DNEk6DlnWkY.png\"\n    alt=\"Liquid AI\"\n    style=\"width: 100%; max-width: 100%; height: auto; display: inline-block; margin-bottom: 0.5em; margin-top: 0.5em;\"\n  />\n</div>\n<div style=\"display: flex; justify-content: center; gap: 0.5em;\">\n<a href=\"https://playground.liquid.ai/\"><strong>Try LFM</strong></a> • <a href=\"https://docs.liquid.ai/lfm/getting-started/welcome\"><strong>Docs</strong></a> • <a href=\"https://leap.liquid.ai/\"><strong>LEAP</strong></a> • <a href=\"https://discord.com/invite/liquid-ai\"><strong>Discord</strong></a>\n</div>\n</center>\n\n<br>\n\n# LFM2-24B-A2B-GGUF\n\nLFM2 is a family of hybrid models designed for on-device deployment. LFM2-24B-A2B is the largest model in the family, scaling the architecture to 24 billion parameters while keeping inference efficient.\n\n- **Best-in-class efficiency**: A 24B MoE model with only 2B active parameters per token, fitting in 32 GB of RAM for deployment on consumer laptops and desktops.\n- **Fast edge inference**: 112 tok/s decode on AMD CPU, 293 tok/s on H100. Fits in 32B GB of RAM with day-one support llama.cpp, vLLM, and SGLang.\n- **Predictable scaling**: Quality improves log-linearly from 350M to 24B total parameters, confirming the LFM2 hybrid architecture scales reliably across nearly two orders of magnitude.\n\n![image](https://cdn-uploads.huggingface.co/production/uploads/61b8e2ba285851687028d395/VhdjtAT5zWTWdVYgW69O0.png)\n\nFind more information about LFM2-24B-A2B in our [blog post](https://www.liquid.ai/blog/).\n\n## How to run LFM2\n\nExample usage with [llama.cpp](https://github.com/ggml-org/llama.cpp):\n\n```\nllama-cli -hf LiquidAI/LFM2-24B-A2B-GGUF\n```\n\n## 📬 Contact\n\n- Got questions or want to connect? [Join our Discord community](https://discord.com/invite/liquid-ai)\n- If you are interested in custom solutions with edge deployment, please contact [our sales team](https://www.liquid.ai/contact).\n",
    "related_quantizations": []
  },
  "tags": [
    "transformers",
    "gguf",
    "liquid",
    "lfm2",
    "edge",
    "moe",
    "llama.cpp",
    "text-generation",
    "en",
    "ar",
    "zh",
    "fr",
    "de",
    "ja",
    "ko",
    "es",
    "pt",
    "base_model:LiquidAI/LFM2-24B-A2B",
    "base_model:quantized:LiquidAI/LFM2-24B-A2B",
    "license:other",
    "endpoints_compatible",
    "region:us",
    "conversational"
  ],
  "likes": 122,
  "downloads": 56748,
  "gated": false,
  "private": false,
  "last_modified": "2026-03-30T12:54:26.000Z",
  "created_at": "2026-02-17T12:53:20.000Z",
  "pipeline_tag": "text-generation",
  "library_name": "transformers"
}
Source payload excerpt (from Hugging Face API)
{
  "_id": "699464c0eb58cf55ba03983f",
  "id": "LiquidAI/LFM2-24B-A2B-GGUF",
  "modelId": "LiquidAI/LFM2-24B-A2B-GGUF",
  "sha": "35b3ac71cf7ddf668d5869843292a8945c399dcd",
  "createdAt": "2026-02-17T12:53:20.000Z",
  "lastModified": "2026-03-30T12:54:26.000Z",
  "author": "LiquidAI",
  "downloads": 56748,
  "likes": 122,
  "gated": false,
  "private": false,
  "pipeline_tag": "text-generation",
  "library_name": "transformers",
  "siblings_count": 10
}