GraySoft
Projects Models Compare Cloud benchmarks FAQ Download guIDE →
Model Intelligence Sheet

nphearum/Opus4.8-tuned-super-fast-GGUF overview

Opus4.8 tuned super fast GGUF : GGUF This model was finetuned and converted to GGUF format Example usage : For text only LLMs: llama cli hf nphearum/Opus4.8 tu…

gguflfm2llama.cppunslothendpoints_compatibleregion:usconversational

Runs locally from ~492.0 MB disk (4 GB VRAM class GPUs with llama.cpp / guIDE).

Downloads
0
Likes
0
Pipeline
Author

Repository Files & Downloads

8 GGUF files detected
Direct downloads for local inference
FileTypeQuantizationSizeLink
LFM2.5-1.2B-Instruct.BF16.ggufGGUFGGUF2.18 GBDownload
LFM2.5-1.2B-Instruct.F16.ggufGGUFGGUF2.18 GBDownload
LFM2.5-1.2B-Instruct.Q2_K_L.ggufGGUFGGUF492.0 MBDownload
LFM2.5-1.2B-Instruct.Q3_K_M.ggufGGUFGGUF572.5 MBDownload
LFM2.5-1.2B-Instruct.Q4_K_M.ggufGGUFGGUF697.0 MBDownload
LFM2.5-1.2B-Instruct.Q5_K_M.ggufGGUFGGUF804.3 MBDownload
LFM2.5-1.2B-Instruct.Q6_K.ggufGGUFGGUF918.2 MBDownload
LFM2.5-1.2B-Instruct.Q8_0.ggufGGUFGGUF1.16 GBDownload

Model Details

Model IDnphearum/Opus4.8-tuned-super-fast-GGUF
Authornphearum
Pipeline
License
Base model
Last modified2026-07-02T16:46:07.000Z

Model README

---

tags:

  • gguf
  • llama.cpp
  • unsloth

---

Opus4.8-tuned-super-fast-GGUF : GGUF

This model was finetuned and converted to GGUF format

Example usage:

  • For text only LLMs: llama-cli -hf nphearum/Opus4.8-tuned-super-fast-GGUF --jinja
  • For multimodal models: llama-mtmd-cli -hf nphearum/Opus4.8-tuned-super-fast-GGUF --jinja

Run nphearum/Opus4.8-tuned-super-fast-GGUF with guIDE

Download guIDE — the AI-native code editor with local LLM inference and 69 built-in tools.

Download guIDE → · Browse 524k+ models · Compare models

Source: Hugging Face · Compare models