GraySoft
Projects Models Compare Cloud benchmarks FAQ Download guIDE →
Model Intelligence Sheet

realoperator42/trinity-nano-stheno-GGUF overview

trinity nano stheno — GGUF realoperator42/trinity nano stheno lora attention only LoRA, r=16, alpha=32 merged into arcee ai/Trinity Nano Preview https://huggin…

ggufllama.cppchatmlroleplaybase_model:arcee-ai/Trinity-Nano-Previewbase_model:quantized:arcee-ai/Trinity-Nano-Previewlicense:apache-2.0endpoints_compatibleregion:usconversational

Runs locally from ~3.50 GB disk (4 GB VRAM class GPUs with llama.cpp / guIDE).

Downloads
0
Likes
0
Pipeline

Repository Files & Downloads

2 GGUF files detected
Direct downloads for local inference
FileTypeQuantizationSizeLink
trinity-nano-stheno-Q4_K_M.ggufGGUFQ4_K_M3.50 GBDownload
trinity-nano-stheno-Q8_0.ggufGGUFQ8_06.08 GBDownload

Model Details

Model IDrealoperator42/trinity-nano-stheno-GGUF
Authorrealoperator42
Pipeline
Licenseapache-2.0
Base modelarcee-ai/Trinity-Nano-Preview
Last modified2026-08-08T22:10:02.000Z

Model README

---

license: apache-2.0

base_model: arcee-ai/Trinity-Nano-Preview

tags: [gguf, llama.cpp, chatml, roleplay]

---

trinity-nano-stheno — GGUF

realoperator42/trinity-nano-stheno-lora (attention-only LoRA, r=16, alpha=32) merged into

arcee-ai/Trinity-Nano-Preview and converted to GGUF.

Architecture is afmoe, so you need a llama.cpp build that includes AFMOE support.

The embedded chat template is the adapter's corrected ChatML, with the training-only

{% generation %} markers removed so --jinja parses cleanly.

llama-cli -m trinity-nano-stheno-Q4_K_M.gguf --jinja -cnv

Files

| File | Size |

| --- | --- |

| trinity-nano-stheno-Q4_K_M.gguf | 3.8 GB |

| trinity-nano-stheno-Q8_0.gguf | 6.5 GB |

Run realoperator42/trinity-nano-stheno-GGUF with guIDE

Download guIDE — the AI-native code editor with local LLM inference and 69 built-in tools.

Download guIDE → · Browse 524k+ models · Compare models

Source: Hugging Face · Compare models