Views
No views yet

Gemma4ForCausalLM and quantization config accordinglyembed_tokens stays BF16, all norms preserved, so we retain all the nvidia/Gemma-4-31B-IT-NVFP4 optimizations.embed_tokens stays BF16, preventing noise from propagating through all layers