GraySoft
Projects Models Compare Cloud benchmarks FAQ Download guIDE →
Model Intelligence Sheet

mparvin/tinyllama-1.1b-chat-v1.0-GGUF overview

tinyllama 1.1b chat v1.0 GGUF GGUF conversions of TinyLlama/TinyLlama 1.1B Chat v1.0 https://huggingface.co/TinyLlama/TinyLlama 1.1B Chat v1.0 . These weights …

ggufquantizedllama.cppbase_model:TinyLlama/TinyLlama-1.1B-Chat-v1.0base_model:quantized:TinyLlama/TinyLlama-1.1B-Chat-v1.0endpoints_compatibleregion:usconversational

Runs locally from ~636.9 MB disk (4 GB VRAM class GPUs with llama.cpp / guIDE).

Downloads
0
Likes
0
Pipeline
Author

Repository Files & Downloads

3 GGUF files detected
Direct downloads for local inference
FileTypeQuantizationSizeLink
tinyllama-1.1b-chat-v1.0-q4_k_m.ggufGGUFQ4_K_M636.9 MBDownload
tinyllama-1.1b-chat-v1.0-q5_k_m.ggufGGUFQ5_K_M745.8 MBDownload
tinyllama-1.1b-chat-v1.0-q8_0.ggufGGUFQ8_01.09 GBDownload

Model Details

Model IDmparvin/tinyllama-1.1b-chat-v1.0-GGUF
Authormparvin
Pipeline
License
Base modelTinyLlama/TinyLlama-1.1B-Chat-v1.0
Last modified2026-09-01T04:45:50.000Z

Model README

---

base_model: TinyLlama/TinyLlama-1.1B-Chat-v1.0

tags:

  • gguf
  • quantized
  • llama.cpp

---

tinyllama-1.1b-chat-v1.0-GGUF

GGUF conversions of TinyLlama/TinyLlama-1.1B-Chat-v1.0.

These weights were converted with llama.cpp and quantized for local / Ollama use. This is not the original model; it is a community conversion of the source weights.

Quantizations

Q4_K_M, Q5_K_M, Q8_0

Source

Run mparvin/tinyllama-1.1b-chat-v1.0-GGUF with guIDE

Download guIDE — the AI-native code editor with local LLM inference and 69 built-in tools.

Download guIDE → · Browse 524k+ models · Compare models

Source: Hugging Face · Compare models