GraySoft
Projects Models Compare Cloud benchmarks FAQ Download guIDE →
Model Intelligence Sheet

Yingyaeliae/grok-oss-Apollyon-8B-heretic-GGUF overview

Grok OSS Apollyon 8B Heretic GGUF Quantizations Educational Purpose Notice This model and its respective quantizations are provided strictly for educational, r…

ggufllama-cppquantizationbase_model:c4tdr0ut/grok-oss-Apollyon-8Bbase_model:quantized:c4tdr0ut/grok-oss-Apollyon-8Blicense:otherendpoints_compatibleregion:usconversational

Runs locally from ~3.74 GB disk (4 GB VRAM class GPUs with llama.cpp / guIDE).

Downloads
0
Likes
0
Pipeline

Repository Files & Downloads

5 GGUF files detected
Direct downloads for local inference
FileTypeQuantizationSizeLink
apollyon-8b-IQ4_NL.ggufGGUFIQ4_NL4.38 GBDownload
apollyon-8b-Q3_K_M.ggufGGUFQ3_K_M3.74 GBDownload
apollyon-8b-Q4_K_M.ggufGGUFQ4_K_M4.58 GBDownload
apollyon-8b-Q5_K_M.ggufGGUFQ5_K_M5.34 GBDownload
apollyon-8b-Q8_0.ggufGGUFQ8_07.95 GBDownload

Model Details

Model IDYingyaeliae/grok-oss-Apollyon-8B-heretic-GGUF
AuthorYingyaeliae
Pipeline
Licenseother
Base modelc4tdr0ut/grok-oss-Apollyon-8B
Last modified2026-07-10T02:05:15.000Z

Model README

---

license: other

base_model: c4tdr0ut/grok-oss-Apollyon-8B

tags:

  • llama-cpp
  • gguf
  • quantization

---

Grok-OSS Apollyon 8B Heretic - GGUF Quantizations

Educational Purpose Notice

This model and its respective quantizations are provided strictly for educational, research, and technical evaluation purposes.

Critical Instructions for Users

  1. Upstream Terms: Users are explicitly required to thoroughly read, understand, and consider the original model's licensing instructions, safety guidelines, and terms of use provided by the creator at Yingyaeliae/grok-oss-Apollyon-8B-heretic.
  2. User Responsibility & Liability: By downloading, hosting, or interacting with this model, the user assumes full and sole responsibility for any and all content generated. The quantizer accepts no liability for misuse, harmful outputs, or secondary deployments of this artifact.

Available Quantizations

Optimized for local CPU/GPU inference via llama.cpp. Use Q5_K_M or IQ4_NL for an optimal balance of throughput speed and accuracy on consumer hardware.

Run Yingyaeliae/grok-oss-Apollyon-8B-heretic-GGUF with guIDE

Download guIDE — the AI-native code editor with local LLM inference and 69 built-in tools.

Download guIDE → · Browse 524k+ models · Compare models

Source: Hugging Face · Compare models