GraySoft
Projects Models Compare Cloud benchmarks FAQ Download guIDE →
Model Intelligence Sheet

UncannyEcho/Aura-Medium-v1.1-GGUF overview

Aura Medium v1.1 GGUF — Portable Local Inference The portable 12B Aura release for local llama.cpp inference, evaluation, and deployment. Aura Medium v1.1 GGUF…

ggufendataset:UncannyEcho/AuraPersonalitydataset:UncannyEcho/AuraAblationbase_model:google/gemma-4-12B-itbase_model:quantized:google/gemma-4-12B-itlicense:apache-2.0endpoints_compatibleregion:usconversational

Runs locally from ~167.0 MB disk (4 GB VRAM class GPUs with llama.cpp / guIDE).

Downloads
392
Likes
0
Pipeline

Repository Files & Downloads

7 GGUF files detected
Direct downloads for local inference
FileTypeQuantizationSizeLink
Aura-Medium-v1p1-F16.ggufGGUFF1622.20 GBDownload
Aura-Medium-v1p1-Q3_K_M.ggufGGUFQ3_K_M5.67 GBDownload
Aura-Medium-v1p1-Q4_K_M.ggufGGUFQ4_K_M6.87 GBDownload
Aura-Medium-v1p1-Q5_K_M.ggufGGUFQ5_K_M7.96 GBDownload
Aura-Medium-v1p1-Q6_K.ggufGGUFQ6_K9.11 GBDownload
Aura-Medium-v1p1-Q8_0.ggufGGUFQ8_011.80 GBDownload
Aura-Medium-v1p1-mmproj-BF16.ggufGGUFBF16167.0 MBDownload

Model Details

Model IDUncannyEcho/Aura-Medium-v1.1-GGUF
AuthorUncannyEcho
Pipeline
Licenseapache-2.0
Base modelgoogle/gemma-4-12B-it
Last modified2026-08-15T20:50:10.000Z

Model README

---

license: apache-2.0

datasets:

  • UncannyEcho/AuraPersonality
  • UncannyEcho/AuraAblation

language:

  • en

base_model:

  • google/gemma-4-12B-it

---

Aura Medium v1.1 GGUF — Portable Local Inference

> The portable 12B Aura release for local llama.cpp inference, evaluation, and deployment.

Aura Medium v1.1 GGUF is the portable local-inference edition of Aura Medium v1.1. It is derived from Google’s Gemma 4 12B IT checkpoint and combines the Aura personality and ablation adapters into one model.

The v1.1 release has seen significant improvements to the personality adapter, merged and applied with the same ablation adapter as v1.

This release is intended for llama.cpp-based inference, evaluation, further research, and deployment across systems with varying memory and performance requirements.

Highlights

  • Parameters: 12 billion
  • Quantizations: Q3_K_M, Q4_K_M, Q5_K_M, Q6_K, Q8_0, and F16
  • Foundation: Google Gemma 4 12B IT
  • Datasets: AuraPersonality and AuraAblation
  • Runtime: llama.cpp and compatible applications
  • Deployment: Local GPU workstations, laptops, and servers
  • Capabilities: Text, vision, audio, and video
  • Focus: Portability, privacy, natural conversation, evaluation, and research

About Aura

Aura is designed for local and on-device deployment across multiple tasks. Aura can serve as a companion or friend, as deemed appropriate by the user, while retaining the broader capabilities of Gemma 4 for:

  • Natural conversation
  • Creative writing
  • Reasoning
  • Image understanding
  • Audio and video understanding
  • Tool use
  • Evaluation and research

Aura Medium v1.1 occupies the middle position in the Aura family, providing greater capability than the compact E4B release while remaining practical for high-end consumer hardware.

Run UncannyEcho/Aura-Medium-v1.1-GGUF with guIDE

Download guIDE — the AI-native code editor with local LLM inference and 69 built-in tools.

Download guIDE → · Browse 524k+ models · Compare models

Source: Hugging Face · Compare models