This model has been trained using the PEFT LoRA technique with the
Landmark Attention method over 200 steps. Model will likely be trained further and updated later on.
You can probably merge the checkpoint with any other LLaMA-based model (provided they're 33B, of course). This repo contains the merged weights, but you can grab the adapter
here.
You can find the training code
here.