GraySoft
Projects Models Compare Cloud benchmarks FAQ Download guIDE →
Model Intelligence Sheet

ggml-org/gpt-oss-120b-GGUF overview

gpt oss 120b Run with https://llama.app bash llama serve hf ggml org/gpt oss 120b GGUF Source models https://huggingface.co/openai/gpt oss 120b https://hugging…

ggufquantizedtext-generationbase_model:openai/gpt-oss-120bbase_model:quantized:openai/gpt-oss-120blicense:apache-2.0endpoints_compatibleregion:usconversational

Runs locally from ~809.8 MB disk (4 GB VRAM class GPUs with llama.cpp / guIDE).

Downloads
17,442
Likes
75
Pipeline
text-generation
Author

Repository Files & Downloads

3 GGUF files detected
Direct downloads for local inference
FileTypeQuantizationSizeLink
eagle3-gpt-oss-120b-BF16.ggufGGUFBF161.48 GBDownload
eagle3-gpt-oss-120b-Q8_0.ggufGGUFQ8_0809.8 MBDownload
gpt-oss-120b-MXFP4.ggufGGUFGGUF59.03 GBDownload

Model Details

Model IDggml-org/gpt-oss-120b-GGUF
Authorggml-org
Pipelinetext-generation
Licenseapache-2.0
Base modelopenai/gpt-oss-120b
Last modified2026-07-28T08:34:41.000Z

Model README

---

license: apache-2.0

pipeline_tag: text-generation

tags:

  • gguf
  • quantized

base_model:

  • openai/gpt-oss-120b

---

gpt-oss-120b

Run with https://llama.app

llama serve -hf ggml-org/gpt-oss-120b-GGUF

Source models

  • https://huggingface.co/openai/gpt-oss-120b
  • https://huggingface.co/nvidia/gpt-oss-120b-Eagle3-v3

TODOs

  • add info

> [!IMPORTANT]

> This model is automatically converted using https://github.com/ggml-org/convert

Run ggml-org/gpt-oss-120b-GGUF with guIDE

Download guIDE — the AI-native code editor with local LLM inference and 69 built-in tools.

Download guIDE → · Browse 524k+ models · Compare models

Source: Hugging Face · Compare models