GraySoft
Projects Models Compare Cloud benchmarks FAQ Download guIDE →
Model Intelligence Sheet

liodon-ai/Ornith-1.0-35B-GGUF-imatrix-GGUF overview

Ornith 1.0 35B GGUF — iMatrix GGUF GGUF quantizations of deepreinforce ai/Ornith 1.0 35B GGUF https://huggingface.co/deepreinforce ai/Ornith 1.0 35B GGUF , pub…

ggufollamalocal-llmllama.cpplm-studioquantizedimatrixsub-4-bittext-generationbase_model:deepreinforce-ai/Ornith-1.0-35B-GGUFbase_model:quantized:deepreinforce-ai/Ornith-1.0-35B-GGUFlicense:otherendpoints_compatibleregion:usconversational

Runs locally from ~6.97 GB disk (8 GB VRAM class GPUs with llama.cpp / guIDE).

Downloads
8,587
Likes
2
Pipeline
text-generation
Author

Repository Files & Downloads

13 GGUF files detected
Direct downloads for local inference
FileTypeQuantizationSizeLink
Ornith-1.0-35B-GGUF-IQ1_S.ggufGGUFIQ1_S6.97 GBDownload
Ornith-1.0-35B-GGUF-IQ2_M.ggufGGUFIQ2_M10.86 GBDownload
Ornith-1.0-35B-GGUF-IQ2_S.ggufGGUFIQ2_S9.92 GBDownload
Ornith-1.0-35B-GGUF-IQ2_XS.ggufGGUFIQ2_XS9.79 GBDownload
Ornith-1.0-35B-GGUF-IQ3_M.ggufGGUFIQ3_M14.38 GBDownload
Ornith-1.0-35B-GGUF-IQ3_XS.ggufGGUFIQ3_XS13.49 GBDownload
Ornith-1.0-35B-GGUF-IQ4_XS.ggufGGUFIQ4_XS17.44 GBDownload
Ornith-1.0-35B-GGUF-Q2_K.ggufGGUFQ2_K12.05 GBDownload
Ornith-1.0-35B-GGUF-Q3_K_M.ggufGGUFQ3_K_M15.61 GBDownload
Ornith-1.0-35B-GGUF-Q4_K_M.ggufGGUFQ4_K_M19.71 GBDownload
Ornith-1.0-35B-GGUF-Q5_K_M.ggufGGUFQ5_K_M23.03 GBDownload
Ornith-1.0-35B-GGUF-Q6_K.ggufGGUFQ6_K26.56 GBDownload
Ornith-1.0-35B-GGUF-Q8_0.ggufGGUFQ8_034.37 GBDownload

Model Details

Model IDliodon-ai/Ornith-1.0-35B-GGUF-imatrix-GGUF
Authorliodon-ai
Pipelinetext-generation
Licenseother
Base modeldeepreinforce-ai/Ornith-1.0-35B-GGUF
Last modified2026-07-13T04:54:06.000Z

Model README

---

license: other

base_model: deepreinforce-ai/Ornith-1.0-35B-GGUF

base_model_relation: quantized

pipeline_tag: text-generation

library_name: gguf

tags:

  • gguf
  • ollama
  • local-llm
  • llama.cpp
  • lm-studio
  • quantized
  • imatrix
  • sub-4-bit

quantized_by: liodon-ai

---

Ornith-1.0-35B-GGUF — iMatrix GGUF

GGUF quantizations of deepreinforce-ai/Ornith-1.0-35B-GGUF, published by Liodon AI.

Quick Start

llama.cpp

llama-cli -hf liodon-ai/Ornith-1.0-35B-GGUF-imatrix-GGUF:Q4_K_M

Ollama

ollama run hf.co/liodon-ai/Ornith-1.0-35B-GGUF-imatrix-GGUF:Q4_K_M

LM Studio / Jan — search liodon-ai/Ornith-1.0-35B-GGUF-imatrix-GGUF and pick your quant.

Quants

| Quant | Size | VRAM est. | Notes |

|-------|------|-----------|-------|

| IQ2_M | 11.66 GB | ~13 GB | 2-bit, iMatrix — smallest usable |

| IQ3_M | 15.44 GB | ~18 GB | 3-bit, iMatrix — great quality/size tradeoff |

| IQ4_XS | 18.73 GB | ~22 GB | 4-bit extra-small, iMatrix |

| Q4_K_M | 21.17 GB | ~24 GB | 4-bit, iMatrix-calibrated (recommended) |

| Q5_K_M | 24.73 GB | ~28 GB | 5-bit, iMatrix-calibrated |

| Q6_K | 28.51 GB | ~33 GB | 6-bit, iMatrix-calibrated, near-lossless |

| Q8_0 | 36.90 GB | ~42 GB | 8-bit, essentially lossless |

What is iMatrix?

Standard quantization treats all weights equally. iMatrix runs 128 calibration chunks through

the full-precision model to find which weights matter most, then allocates more precision where

it counts. At Q2/Q3/Q4 this means noticeably better coherence and instruction-following —

same file size, better output.

Calibration: 2M tokens of WikiText-103.

> Also see plain (non-iMatrix) quants: liodon-ai/Ornith-1.0-35B-GGUF-GGUF

Source

---

Quantized by Liodon AI

Run liodon-ai/Ornith-1.0-35B-GGUF-imatrix-GGUF with guIDE

Download guIDE — the AI-native code editor with local LLM inference and 69 built-in tools.

Download guIDE → · Browse 524k+ models · Compare models

Source: Hugging Face · Compare models