EXPERIMENTAL FIX... THE MODEL DOESNT HAVE ANY ISSUE... THOUGH IT DOES SEPPOKU... one of the issue we find... not really a massive issue!
just follow our recommended setting... this is just a remnant
This llama model was trained 2x faster with
Unsloth and Huggingface's TRL library.