GraySoft
Projects Models Compare Cloud benchmarks FAQ Download guIDE →
Model Intelligence Sheet

UncannyEcho/Aura-Medium-v1-GGUF overview

Aura Medium v1 GGUF — Portable Local Inference The portable 12B Aura release for local llama.cpp inference, evaluation, and deployment. Aura Medium v1 GGUF is …

ggufendataset:UncannyEcho/AuraPersonalitydataset:UncannyEcho/AuraAblationbase_model:google/gemma-4-12B-itbase_model:quantized:google/gemma-4-12B-itlicense:apache-2.0endpoints_compatibleregion:usconversational

Runs locally from ~167.0 MB disk (4 GB VRAM class GPUs with llama.cpp / guIDE).

Downloads
276
Likes
1
Pipeline

Repository Files & Downloads

7 GGUF files detected
Direct downloads for local inference
FileTypeQuantizationSizeLink
Aura-12B-F16.ggufGGUFF1622.20 GBDownload
Aura-12B-Q3_K_M.ggufGGUFQ3_K_M5.67 GBDownload
Aura-12B-Q4_K_M.ggufGGUFQ4_K_M6.87 GBDownload
Aura-12B-Q5_K_M.ggufGGUFQ5_K_M7.96 GBDownload
Aura-12B-Q6_K.ggufGGUFQ6_K9.11 GBDownload
Aura-12B-Q8_0.ggufGGUFQ8_011.80 GBDownload
Aura-12B-mmproj-BF16.ggufGGUFBF16167.0 MBDownload

Model Details

Model IDUncannyEcho/Aura-Medium-v1-GGUF
AuthorUncannyEcho
Pipeline
Licenseapache-2.0
Base modelgoogle/gemma-4-12B-it
Last modified2026-08-15T20:50:03.000Z

Model README

---

license: apache-2.0

datasets:

  • UncannyEcho/AuraPersonality
  • UncannyEcho/AuraAblation

language:

  • en

base_model:

  • google/gemma-4-12B-it

---

Aura Medium v1 GGUF — Portable Local Inference

> The portable 12B Aura release for local llama.cpp inference, evaluation, and deployment.

Aura Medium v1 GGUF is the portable local-inference edition of Aura Medium v1. It is derived from Google’s Gemma 4 12B IT checkpoint and combines the Aura personality and ablation adapters into one model.

This release is intended for llama.cpp-based inference, evaluation, further research, and deployment across systems with varying memory and performance requirements.

Highlights

  • Parameters: 12 billion
  • Quantizations: Q3_K_M, Q4_K_M, Q5_K_M, Q6_K, Q8_0, and F16
  • Foundation: Google Gemma 4 12B IT
  • Datasets: AuraPersonality and AuraAblation
  • Runtime: llama.cpp and compatible applications
  • Deployment: Local GPU workstations, laptops, and servers
  • Capabilities: Text, vision, audio, and video
  • Focus: Portability, privacy, natural conversation, evaluation, and research

About Aura

Aura is designed for local and on-device deployment across multiple tasks. Aura can serve as a companion or friend, as deemed appropriate by the user, while retaining the broader capabilities of Gemma 4 for:

  • Natural conversation
  • Creative writing
  • Reasoning
  • Image understanding
  • Audio and video understanding
  • Tool use
  • Evaluation and research

Aura Medium v1 occupies the middle position in the Aura family, providing greater capability than the compact E4B release while remaining practical for high-end consumer hardware.

Run UncannyEcho/Aura-Medium-v1-GGUF with guIDE

Download guIDE — the AI-native code editor with local LLM inference and 69 built-in tools.

Download guIDE → · Browse 524k+ models · Compare models

Source: Hugging Face · Compare models