GraySoft
Projects Models Compare Cloud benchmarks FAQ Download guIDE →
Model Intelligence Sheet

ggml-org/SmolLM2-135M-GGUF overview

SmolLM2 135M Run with https://llama.app bash llama serve hf ggml org/SmolLM2 135M GGUF Source models https://huggingface.co/HuggingFaceTB/SmolLM2 135M IMPORTAN…

ggufquantizedtext-generationbase_model:HuggingFaceTB/SmolLM2-135Mbase_model:quantized:HuggingFaceTB/SmolLM2-135Mlicense:apache-2.0endpoints_compatibleregion:us

Runs locally from ~96.3 MB disk (4 GB VRAM class GPUs with llama.cpp / guIDE).

Downloads
0
Likes
0
Pipeline
text-generation
Author

Repository Files & Downloads

3 GGUF files detected
Direct downloads for local inference
FileTypeQuantizationSizeLink
SmolLM2-135M-BF16.ggufGGUFBF16258.3 MBDownload
SmolLM2-135M-Q4_K_M.ggufGGUFQ4_K_M96.3 MBDownload
SmolLM2-135M-Q8_0.ggufGGUFQ8_0138.1 MBDownload

Model Details

Model IDggml-org/SmolLM2-135M-GGUF
Authorggml-org
Pipelinetext-generation
Licenseapache-2.0
Base modelHuggingFaceTB/SmolLM2-135M
Last modified2026-08-23T13:45:00.000Z

Model README

---

license: apache-2.0

pipeline_tag: text-generation

tags:

  • gguf
  • quantized

base_model:

  • HuggingFaceTB/SmolLM2-135M

---

SmolLM2-135M

Run with https://llama.app

llama serve -hf ggml-org/SmolLM2-135M-GGUF

Source models

  • https://huggingface.co/HuggingFaceTB/SmolLM2-135M

> [!IMPORTANT]

> This model is automatically converted using https://github.com/ggml-org/convert

Run ggml-org/SmolLM2-135M-GGUF with guIDE

Download guIDE — the AI-native code editor with local LLM inference and 69 built-in tools.

Download guIDE → · Browse 524k+ models · Compare models

Source: Hugging Face · Compare models