UncannyEcho/Aura-v1.1-GGUF overview
Aura v1.1 GGUF — Portable Local Inference A portable, private, and flexible edition of Aura for local inference. Aura v1.1 GGUF is based on the QAT derived Gem…
Runs locally from ~944.5 MB disk (4 GB VRAM class GPUs with llama.cpp / guIDE).
Repository Files & Downloads
Model Details
| Model ID | UncannyEcho/Aura-v1.1-GGUF |
|---|---|
| Author | UncannyEcho |
| Pipeline | — |
| License | apache-2.0 |
| Base model | google/gemma-4-E4B-it |
| Last modified | 2026-08-15T20:50:24.000Z |
Model README
---
license: apache-2.0
datasets:
- UncannyEcho/AuraPersonality
- UncannyEcho/AuraAblation
language:
- en
base_model:
- google/gemma-4-E4B-it
---
Aura v1.1 GGUF — Portable Local Inference
> A portable, private, and flexible edition of Aura for local inference.
Aura v1.1 GGUF is based on the QAT-derived Gemma 4 E4B Q4_0 model and uses the same canonical combined Aura adapter as the other Aura v1.1 releases.
The v1.1 release has seen significant improvements to the personality adapter, merged and applied with the same ablation adapter as v1.
Designed for the llama.cpp ecosystem, this edition supports local deployment across desktops, laptops, and capable edge devices.
Highlights
- Format: GGUF
- Foundation: QAT-derived Gemma 4 E4B Q4_0
- Runtime:
llama.cppand compatible applications - Deployment: Local and on-device
- Focus: Portability, privacy, and broad hardware compatibility
About Aura
Aura is designed for on-device deployment across multiple tasks. Aura can be a companion or friend, as deemed necessary by the user, or serve as a flexible private assistant for:
- Natural conversation
- Creative writing
- Reasoning
- Multimodal interaction
- Private and offline workflows
Run UncannyEcho/Aura-v1.1-GGUF with guIDE
Download guIDE — the AI-native code editor with local LLM inference and 69 built-in tools.
Source: Hugging Face · Compare models