This fine-tuning dataset contains 260 legal classification tasks derived from the Supreme Court and Songer Court of Appeals databases, totalling over 500k training examples and 2B tokens. This dataset was used to train Lawma 8B and Lawma 70B. The Lawma models outperform GPT-4 on 95% of these legal tasks, on average by over 17 accuracy points. See our arXiv preprint and GitHub repository for more details.
Our reasons to study these legal classification tasks… See the full description on the dataset page:
https://huggingface.co/datasets/ricdomolm/lawma-all-tasks.