GraySoft
Projects Models Compare Cloud benchmarks FAQ Download guIDE →
Model Intelligence Sheet

Avtrkrb/granite-claude-h-350m-GGUF overview

granite claude h 350m GGUF GGUF quantizations of: Avtrkrb/granite claude h 350m These files are intended for inference using: llama.cpp LM Studio Open WebUI Ja…

ggufgranitellama-cppreasoningquantizedlocal-llmtext-generationenbase_model:Avtrkrb/granite-claude-h-350mbase_model:quantized:Avtrkrb/granite-claude-h-350mlicense:apache-2.0endpoints_compatibleregion:usconversational

Runs locally from ~247.4 MB disk (4 GB VRAM class GPUs with llama.cpp / guIDE).

Downloads
0
Likes
0
Pipeline
text-generation
Author

Repository Files & Downloads

6 GGUF files detected
Direct downloads for local inference
FileTypeQuantizationSizeLink
granite-claude-h-350m-F16.ggufGGUFF16800.2 MBDownload
granite-claude-h-350m-Q4_0.ggufGGUFQ4_0247.4 MBDownload
granite-claude-h-350m-Q4_K_M.ggufGGUFQ4_K_M253.7 MBDownload
granite-claude-h-350m-Q5_K_M.ggufGGUFQ5_K_M291.2 MBDownload
granite-claude-h-350m-Q6_K.ggufGGUFQ6_K331.0 MBDownload
granite-claude-h-350m-Q8_0.ggufGGUFQ8_0427.3 MBDownload

Model Details

Model IDAvtrkrb/granite-claude-h-350m-GGUF
AuthorAvtrkrb
Pipelinetext-generation
Licenseapache-2.0
Base modelAvtrkrb/granite-claude-h-350m
Last modified2026-06-11T02:21:18.000Z

Model README

---

license: apache-2.0

language:

- en

pipeline_tag: text-generation

tags:

- granite

- gguf

- llama-cpp

- reasoning

- quantized

- local-llm

base_model: Avtrkrb/granite-claude-h-350m

library_name: gguf

---

granite-claude-h-350m-GGUF

GGUF quantizations of:

Avtrkrb/granite-claude-h-350m

These files are intended for inference using:

  • llama.cpp
  • LM Studio
  • Open WebUI
  • Jan
  • KoboldCpp
  • GPT4All
  • Ollama (after conversion/import)

---

Available Quantizations

Typical variants included:

| Quant | Use Case |

|---------|---------|

| Q4_K_M | Best size / quality balance |

| Q5_K_M | Higher quality |

| Q6_K | Near-lossless for most use cases |

| Q8_0 | Highest quality quantized version |

---

Source Model

Merged model:

https://huggingface.co/Avtrkrb/granite-claude-h-350m

Dataset:

https://huggingface.co/datasets/Avtrkrb/combined-reasoning-claude

---

Example llama.cpp Usage

./llama-cli \
  -m granite-claude-h-350m-Q4_K_M.gguf \
  -p "Explain quantum tunneling."

---

Recommended Quant

For most users:

Q4_K_M

offers the best balance between:

  • quality
  • speed
  • memory usage

---

License

This repository follows the licensing terms of the original Granite model.

Run Avtrkrb/granite-claude-h-350m-GGUF with guIDE

Download guIDE — the AI-native code editor with local LLM inference and 69 built-in tools.

Download guIDE → · Browse 524k+ models · Compare models

Source: Hugging Face · Compare models