GraySoft
Projects Models Compare Cloud benchmarks FAQ Download guIDE โ†’
Model Intelligence Sheet

Maxilicious20/Aether-2.1-GGUF overview

Aether 2.1 GGUF Pre quantized GGUF binaries for Aether 2.1 . Trained with SFT Supervised Fine Tuning and PEFT LoRA on top of Qwen2.5 1.5B Instruct , Aether 2.1โ€ฆ

ggufllama.cpplm-studioaethergermanenglishtext-generationdeenbase_model:Qwen/Qwen2.5-1.5B-Instructbase_model:quantized:Qwen/Qwen2.5-1.5B-Instructlicense:apache-2.0endpoints_compatibleregion:usconversational

Runs locally from ~5.7 MB disk (4 GB VRAM class GPUs with llama.cpp / guIDE).

Downloads
0
Likes
0
Pipeline
text-generation

Repository Files & Downloads

3 GGUF files detected
Direct downloads for local inference
FileTypeQuantizationSizeLink
aether_2_1_fp16.ggufGGUFGGUF5.7 MBDownload
aether_2_1_q4_k_m.ggufGGUFQ4_K_M5.7 MBDownload
aether_2_1_q8_0.ggufGGUFQ8_05.7 MBDownload

Model Details

Model IDMaxilicious20/Aether-2.1-GGUF
AuthorMaxilicious20
Pipelinetext-generation
Licenseapache-2.0
Base modelQwen/Qwen2.5-1.5B-Instruct
Last modified2026-08-01T10:24:21.000Z

Model README

---

library_name: gguf

tags:

  • gguf
  • llama.cpp
  • lm-studio
  • aether
  • german
  • english
  • text-generation

license: apache-2.0

language:

  • de
  • en

base_model: Qwen/Qwen2.5-1.5B-Instruct

---

Aether 2.1 - GGUF

Pre-quantized GGUF binaries for Aether 2.1.

Trained with SFT (Supervised Fine-Tuning) and PEFT (LoRA) on top of Qwen2.5-1.5B-Instruct, Aether 2.1 provides lightweight and efficient conversational AI performance in German and English.

> ๐Ÿ”— Looking for the Base / LoRA Adapter?

> If you want to use the Hugging Face Transformers PEFT adapter instead, check out the main repository:

> ๐Ÿ‘‰ Maxilicious20/Aether-2.1

---

๐Ÿ“ฆ Available Files & Quantizations

Choose the right file depending on your system's VRAM/RAM and performance needs:

| Filename | Quantization | Quality | Size | Description / Recommendation |

| :--- | :--- | :--- | :--- | :--- |

| aether_2_1_fp16.gguf | FP16 / F16 | Maximum | ~5.65 MB | Uncompressed full precision adapter GGUF. |

| aether_2_1_q8_0.gguf | Q8_0 | Very High | ~5.65 MB | High-precision 8-bit quantized build. |

| aether_2_1_q4_k_m.gguf | Q4_K_M | Balanced | ~5.65 MB | Recommended. Balanced quantization for ultra-fast local execution. |

---

๐Ÿš€ How to Run Locally

1. LM Studio

  1. Open LM Studio.
  2. Search for Maxilicious20/Aether-2.1-GGUF or paste the repository ID.
  3. Download your preferred quantization (e.g., aether_2_1_q4_k_m.gguf).
  4. Load the model and start chatting!

2. Ollama / llama.cpp

You can run the GGUF file directly using llama.cpp:

./llama-cli -m aether_2_1_q4_k_m.gguf -p "Hello Aether!" -n 256

Run Maxilicious20/Aether-2.1-GGUF with guIDE

Download guIDE โ€” the AI-native code editor with local LLM inference and 69 built-in tools.

Download guIDE โ†’ ยท Browse 524k+ models ยท Compare models

Source: Hugging Face ยท Compare models