-
Introduction
This model was created using Quark Quantization for the Decoder, followed by OGA Model Builder, and finalized with post-processing for NPU deployment.
-
Quantization Strategy
- AWQ / Group 128 / Asymmetric / BFP16 activations / UINT4 weights
-
Base model info:
-
Please refer to
Gemma-3-4b-it for base model info.
Modifications copyright(c) 2024 Advanced Micro Devices,Inc. All rights reserved.