Polish SCOTUS-Dom is a long-document legal text classification dataset
derived from the publicly available
SCOTUS dataset.
The dataset contains Polish translations of U.S. Supreme Court opinions.
The task is an 11-class document classification problem in which the goal
is to predict the legal issue area of a court case.
SCOTUS-Dom is part of the LongContext benchmark introduced with
Polish ModernBERT.
Split
Examples… See the full description on the dataset page:
https://huggingface.co/datasets/mmichall/SCOTUS-Dom.