GraySoft
Projects Models Compare Cloud benchmarks FAQ Download guIDE →
Model Intelligence Sheet

UncannyEcho/Aura-Large-v1-GGUF overview

Aura Large v1 GGUF — Portable Local Inference The portable 31B Aura release for local llama.cpp inference, evaluation, and deployment. Aura Large v1 GGUF is th…

ggufendataset:UncannyEcho/AuraPersonalitydataset:UncannyEcho/AuraAblationbase_model:google/gemma-4-31B-itbase_model:quantized:google/gemma-4-31B-itlicense:apache-2.0endpoints_compatibleregion:usconversational

Runs locally from ~1.12 GB disk (4 GB VRAM class GPUs with llama.cpp / guIDE).

Downloads
444
Likes
1
Pipeline

Repository Files & Downloads

7 GGUF files detected
Direct downloads for local inference
FileTypeQuantizationSizeLink
Aura-31B-F16.ggufGGUFF1657.20 GBDownload
Aura-31B-Q3_K_M.ggufGGUFQ3_K_M14.24 GBDownload
Aura-31B-Q4_K_M.ggufGGUFQ4_K_M17.40 GBDownload
Aura-31B-Q5_K_M.ggufGGUFQ5_K_M20.35 GBDownload
Aura-31B-Q6_K.ggufGGUFQ6_K23.47 GBDownload
Aura-31B-Q8_0.ggufGGUFQ8_030.39 GBDownload
Aura-31B-mmproj-BF16.ggufGGUFBF161.12 GBDownload

Model Details

Model IDUncannyEcho/Aura-Large-v1-GGUF
AuthorUncannyEcho
Pipeline
Licenseapache-2.0
Base modelgoogle/gemma-4-31B-it
Last modified2026-08-15T20:49:51.000Z

Model README

---

license: apache-2.0

datasets:

  • UncannyEcho/AuraPersonality
  • UncannyEcho/AuraAblation

language:

  • en

base_model:

  • google/gemma-4-31B-it

---

Aura Large v1 GGUF — Portable Local Inference

> The portable 31B Aura release for local llama.cpp inference, evaluation, and deployment.

Aura Large v1 GGUF is the portable local-inference edition of Aura Large v1. It is derived from Google’s Gemma 4 31B IT checkpoint and combines the Aura personality and ablation adapters into one model.

This release is intended for llama.cpp-based inference, evaluation, further research, and deployment across systems with varying memory and performance requirements.

Highlights

  • Parameters: 31 billion
  • Quantizations: Q3_K_M, Q4_K_M, Q5_K_M, Q6_K, Q8_0, and F16
  • Foundation: Google Gemma 4 31B IT
  • Datasets: AuraPersonality and AuraAblation
  • Runtime: llama.cpp and compatible applications
  • Deployment: Local GPU workstations, laptops, and servers
  • Capabilities: Text, vision, and video
  • Focus: Portability, privacy, natural conversation, evaluation, and research

About Aura

Aura is designed for local and on-device deployment across multiple tasks. Aura can serve as a companion or friend, as deemed appropriate by the user, while retaining the broader capabilities of Gemma 4 for:

  • Natural conversation
  • Creative writing
  • Reasoning
  • Image understanding
  • Video understanding
  • Tool use
  • Evaluation and research

Aura Large v1 occupies the largest position in the Aura family, providing greater capability than the Medium release while remaining practical for high-end consumer hardware.

Run UncannyEcho/Aura-Large-v1-GGUF with guIDE

Download guIDE — the AI-native code editor with local LLM inference and 69 built-in tools.

Download guIDE → · Browse 524k+ models · Compare models

Source: Hugging Face · Compare models