The high-fidelity Aura release for Transformers inference, evaluation, and research.
Aura v1.1 BF16 is the half-precision edition of Aura, derived from Google’s unquantized Gemma 4 E4B QAT checkpoint and the same canonical combined Aura adapter used across the Aura v1.1 family.
The v1.1 release has seen significant improvements to the personality adapter, merged and applied with the same ablation adapter as v1.
This release is intended for Transformers-based inference, evaluation, further research, and deployment on GPUs with sufficient memory.
Aura is designed for on-device deployment across multiple tasks. Aura can be a companion or friend, as deemed necessary by the user, while retaining the broader capabilities of Gemma 4 for: