License: Must comply with license of Llama2 since it's a model derived from Llama2.
Sheared-LLaMA-1.3B-Pruned is the model pruned from meta-llama/Llama-2-7b-hfwithout continued pre-training.
We used roughly 0.4B tokens to perform the pruning experiment. This model could be a good use to study
effective data mixtures for continued pre-training
comparisons to other pruning techniques
extensive evaluations to understand how pruning affects knowledge and reasoning capabilities of LLMs