GGUF
English
conversational
How to use from
Docker Model Runner
docker model run hf.co/UncannyEcho/Aura-v1-GGUF:
Quick Links

Aura v1 GGUF — Portable Local Inference

A portable, private, and flexible edition of Aura for local inference.

Aura v1 GGUF is based on the QAT-derived Gemma 4 E4B Q4_0 model and uses the same canonical combined Aura adapter as the other Aura v1 releases.

Designed for the llama.cpp ecosystem, this edition supports local deployment across desktops, laptops, and capable edge devices.

Highlights

  • Format: GGUF
  • Foundation: QAT-derived Gemma 4 E4B Q4_0
  • Runtime: llama.cpp and compatible applications
  • Deployment: Local and on-device
  • Focus: Portability, privacy, and broad hardware compatibility

About Aura

Aura is designed for on-device deployment across multiple tasks. Aura can be a companion or friend, as deemed necessary by the user, or serve as a flexible private assistant for:

  • Natural conversation
  • Creative writing
  • Reasoning
  • Multimodal interaction
  • Private and offline workflows
Downloads last month
85
GGUF
Model size
7B params
Architecture
gemma4
Hardware compatibility
Log In to add your hardware

4-bit

8-bit

16-bit

Inference Providers NEW
This model isn't deployed by any Inference Provider. 🙋 Ask for provider support

Model tree for UncannyEcho/Aura-v1-GGUF

Quantized
(341)
this model

Datasets used to train UncannyEcho/Aura-v1-GGUF

Collection including UncannyEcho/Aura-v1-GGUF