UncannyEcho/Aura-v1-GGUF overview
Aura v1 GGUF — Portable Local Inference A portable, private, and flexible edition of Aura for local inference. Aura v1 GGUF is based on the QAT derived Gemma 4…
Runs locally from ~944.5 MB disk (4 GB VRAM class GPUs with llama.cpp / guIDE).
Repository Files & Downloads
Model Details
| Model ID | UncannyEcho/Aura-v1-GGUF |
|---|---|
| Author | UncannyEcho |
| Pipeline | — |
| License | apache-2.0 |
| Base model | google/gemma-4-E4B-it |
| Last modified | 2026-08-15T20:50:17.000Z |
Model README
---
license: apache-2.0
datasets:
- UncannyEcho/AuraPersonality
- UncannyEcho/AuraAblation
language:
- en
base_model:
- google/gemma-4-E4B-it
---
Aura v1 GGUF — Portable Local Inference
> A portable, private, and flexible edition of Aura for local inference.
Aura v1 GGUF is based on the QAT-derived Gemma 4 E4B Q4_0 model and uses the same canonical combined Aura adapter as the other Aura v1 releases.
Designed for the llama.cpp ecosystem, this edition supports local deployment across desktops, laptops, and capable edge devices.
Highlights
- Format: GGUF
- Foundation: QAT-derived Gemma 4 E4B Q4_0
- Runtime:
llama.cppand compatible applications - Deployment: Local and on-device
- Focus: Portability, privacy, and broad hardware compatibility
About Aura
Aura is designed for on-device deployment across multiple tasks. Aura can be a companion or friend, as deemed necessary by the user, or serve as a flexible private assistant for:
- Natural conversation
- Creative writing
- Reasoning
- Multimodal interaction
- Private and offline workflows
Run UncannyEcho/Aura-v1-GGUF with guIDE
Download guIDE — the AI-native code editor with local LLM inference and 69 built-in tools.
Source: Hugging Face · Compare models