GraySoft
Projects Models Compare Cloud benchmarks FAQ Download guIDE →
Model Intelligence Sheet

ggml-org/Llama-2-7B-GGUF overview

Llama 2 7B This model is used mainly for data collection for https://github.com/ggml org/llama.cpp/discussions/4167 Source models https://huggingface.co/meta l…

ggufquantizedtext-generationbase_model:meta-llama/Llama-2-7b-hfbase_model:quantized:meta-llama/Llama-2-7b-hflicense:llama2endpoints_compatibleregion:us

Runs locally from ~3.53 GB disk (4 GB VRAM class GPUs with llama.cpp / guIDE).

Downloads
0
Likes
0
Pipeline
text-generation
Author

Repository Files & Downloads

3 GGUF files detected
Direct downloads for local inference
FileTypeQuantizationSizeLink
Llama-2-7B-F16.ggufGGUFF1612.55 GBDownload
Llama-2-7B-Q4_0.ggufGGUFQ4_03.53 GBDownload
Llama-2-7B-Q8_0.ggufGGUFQ8_06.67 GBDownload

Model Details

Model IDggml-org/Llama-2-7B-GGUF
Authorggml-org
Pipelinetext-generation
Licensellama2
Base modelmeta-llama/Llama-2-7b-hf
Last modified2026-08-25T14:05:55.000Z

Model README

---

license: llama2

pipeline_tag: text-generation

tags:

  • gguf
  • quantized

base_model:

  • meta-llama/Llama-2-7b-hf

---

Llama-2-7B

This model is used mainly for data collection for https://github.com/ggml-org/llama.cpp/discussions/4167

Source models

  • https://huggingface.co/meta-llama/Llama-2-7b-hf

> [!IMPORTANT]

> This model is automatically converted using https://github.com/ggml-org/convert

Run ggml-org/Llama-2-7B-GGUF with guIDE

Download guIDE — the AI-native code editor with local LLM inference and 69 built-in tools.

Download guIDE → · Browse 524k+ models · Compare models

Source: Hugging Face · Compare models