Gemma-3-270M-IT is a ~270-million-parameter instruction-tuned, dense Transformer decoder-only language model developed by Google, part of the Gemma 3 family. It features
GeGLU activation,
Grouped Query Attention (GQA), and
32K native context length. Pre-trained on diverse web-scale corpora and aligned via instruction tuning + RLHF. This distribution is provided by
Aria Compute as an
aria-quant-bundle — a quantized package using
Hadamard pre-processing + per-channel quantization. Optimized for
CPU-only, on-device inference on mobile phones, edge devices, and single-board computers via the
Aria Engine runtime. No GPU or cloud connection is required.
Authenticated dashboard users can download the bundle via:
https://ariacompute.com/dashboard/models
Users (both direct and downstream) should be made aware of the above risks, biases, limitations, and constraints of the model. We recommend: