Training used LoRA rank 128 and alpha set to 32. Context length was set to 16384. But the there's more data in 8k context length so using 8k context length will likely perform better.
Original dataset (before filtering/cleaning):
NilanE/ParallelFiction-Ja_En-100k
This gemma3 model was trained 2x faster with
Unsloth and Huggingface's TRL library.