This dataset contains the metadata of case law opinions used for training and evaluating the Free Law Project Semantic Search Project.
The dataset is curated by Free Law Project by randomly sampling ~1K cases across various courts and jurisdictions from the CourtListener database.
This dataset contains all the metadata associated with each opinion, with opinion_id as the unique identifier. The train split… See the full description on the dataset page:
https://huggingface.co/datasets/freelawproject/opinions-metadata.