Model Intelligence Sheet
mparvin/tinyllama-1.1b-chat-v1.0-GGUF overview
tinyllama 1.1b chat v1.0 GGUF GGUF conversions of TinyLlama/TinyLlama 1.1B Chat v1.0 https://huggingface.co/TinyLlama/TinyLlama 1.1B Chat v1.0 . These weights …
Runs locally from ~636.9 MB disk (4 GB VRAM class GPUs with llama.cpp / guIDE).
Repository Files & Downloads
Model Details
Model README
---
base_model: TinyLlama/TinyLlama-1.1B-Chat-v1.0
tags:
- gguf
- quantized
- llama.cpp
---
tinyllama-1.1b-chat-v1.0-GGUF
GGUF conversions of TinyLlama/TinyLlama-1.1B-Chat-v1.0.
These weights were converted with llama.cpp and quantized for local / Ollama use. This is not the original model; it is a community conversion of the source weights.
Quantizations
Q4_K_M, Q5_K_M, Q8_0
Source
- Base model: TinyLlama/TinyLlama-1.1B-Chat-v1.0
- Converted by HF-2-GGUF
Run mparvin/tinyllama-1.1b-chat-v1.0-GGUF with guIDE
Download guIDE — the AI-native code editor with local LLM inference and 69 built-in tools.
Source: Hugging Face · Compare models