UncannyEcho/Aura-Gaiden-Medium-v1.1-GGUF overview
Aura Gaiden Medium v1.1 GGUF — Portable Local Inference The portable 26B A4B Aura release for local llama.cpp inference, evaluation, and deployment. Aura Gaide…
Runs locally from ~1.11 GB disk (4 GB VRAM class GPUs with llama.cpp / guIDE).
Repository Files & Downloads
| File | Type | Quantization | Size | Link |
|---|---|---|---|---|
| Aura-Gaiden-Medium-v1p1-F16.gguf | GGUF | F16 | 47.04 GB | Download |
| Aura-Gaiden-Medium-v1p1-Q3_K_M.gguf | GGUF | Q3_K_M | 12.37 GB | Download |
| Aura-Gaiden-Medium-v1p1-Q4_K_M.gguf | GGUF | Q4_K_M | 15.64 GB | Download |
| Aura-Gaiden-Medium-v1p1-Q5_K_M.gguf | GGUF | Q5_K_M | 17.82 GB | Download |
| Aura-Gaiden-Medium-v1p1-Q6_K.gguf | GGUF | Q6_K | 21.08 GB | Download |
| Aura-Gaiden-Medium-v1p1-Q8_0.gguf | GGUF | Q8_0 | 25.02 GB | Download |
| mmproj-Aura-Gaiden-Medium-v1p1-BF16.gguf | GGUF | BF16 | 1.11 GB | Download |
Model Details
| Model ID | UncannyEcho/Aura-Gaiden-Medium-v1.1-GGUF |
|---|---|
| Author | UncannyEcho |
| Pipeline | — |
| License | apache-2.0 |
| Base model | google/gemma-4-26B-A4B-it |
| Last modified | 2026-08-15T20:49:44.000Z |
Model README
---
base_model: google/gemma-4-26B-A4B-it
license: apache-2.0
language: en
datasets:
- UncannyEcho/AuraAblation100
- UncannyEcho/AuraPersonality
---
Aura Gaiden Medium v1.1 GGUF — Portable Local Inference
> The portable 26B A4B Aura release for local llama.cpp inference, evaluation, and deployment.
Aura Gaiden Medium v1.1 GGUF is the portable local-inference edition of Aura Gaiden Medium v1.1. It is derived from Google’s Gemma 4 26B A4B IT checkpoint and combines the Aura personality and ablation adapters into one model.
The v1.1 release has seen significant improvements to the personality adapter, merged and applied with the same ablation adapter as v1.
This release is intended for llama.cpp-based inference, evaluation, further research, and deployment across systems with varying memory and performance requirements.
Highlights
- Parameters: 26 billion total, 4 billion active
- Precision: BFloat16
- Foundation: Google Gemma 4 26B A4B IT
- Datasets: AuraPersonality and AuraAblation
- Runtime: Hugging Face Transformers
- Deployment: Local GPU workstations and servers
- Capabilities: Text, vision, audio, and video
- Focus: Fidelity, natural conversation, evaluation, and research
About Aura
Aura is designed for local and on-device deployment across multiple tasks. Aura can serve as a companion or friend, as deemed appropriate by the user, while retaining the broader capabilities of Gemma 4 for:
- Natural conversation
- Creative writing
- Reasoning
- Image understanding
- Audio and video understanding
- Tool use
- Evaluation and research
The Aura Gaiden category is a unique one off experiment. There may or may not be more models in this class.
Run UncannyEcho/Aura-Gaiden-Medium-v1.1-GGUF with guIDE
Download guIDE — the AI-native code editor with local LLM inference and 69 built-in tools.
Source: Hugging Face · Compare models